P
PaperPicks
Conferences
Diyi Yang
65 papers at tracked venues · 46 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
ACL
×18
EMNLP
×11
NeurIPS
×9
CHI
×7
ICLR
×6
CSCW
×4
NAACL
×3
UIST
×3
ICML
×2
AAAI
×1
EACL
×1
Frequent coauthors
Dora Zhao
DBLP profile ↗
ORCID search ↗
×4
Omar Shaikh
DBLP profile ↗
ORCID search ↗
×4
Ryan Louie
DBLP profile ↗
ORCID search ↗
×2
Hua Shen
DBLP profile ↗
ORCID search ↗
×2
Chenglei Si
DBLP profile ↗
ORCID search ↗
×2
Caleb Ziems
DBLP profile ↗
ORCID search ↗
×2
Jing Huang
DBLP profile ↗
ORCID search ↗
×2
Minzhi Li
DBLP profile ↗
ORCID search ↗
×2
Binwei Yao
DBLP profile ↗
ORCID search ↗
×2
John Yang
DBLP profile ↗
ORCID search ↗
×2
Michael J. Ryan
DBLP profile ↗
ORCID search ↗
×2
Potsawee Manakul
DBLP profile ↗
ORCID search ↗
×1
Papers
AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation
EACL 2026
·
Potsawee Manakul
DBLP profile ↗
ORCID search ↗
Can LLM-Simulated Practice and Feedback Upskill Human Counselors? A Randomized Study with 90+ Novice Counselors
CHI 2026
·
Ryan Louie
DBLP profile ↗
ORCID search ↗
Future of Work in the Age of LLMs
ACL 2026
·
Zora Zhiruo Wang
DBLP profile ↗
ORCID search ↗
Generative Interfaces for Language Models
ACL 2026
·
Jiaqi Chen
DBLP profile ↗
ORCID search ↗
Just-In-Time Objectives: A General Approach for Specialized AI Interactions
CHI 2026
·
Michelle S. Lam
DBLP profile ↗
ORCID search ↗
Mapping the Spiral of Silence: Surveying Unspoken Opinions in Online Communities
CHI 2026
·
Dora Zhao
DBLP profile ↗
ORCID search ↗
Putting HUMANS first: Efficient LAM Evaluation with Human Preference Alignment
ACL 2026
·
Woody Haosheng Gan
DBLP profile ↗
ORCID search ↗
Verbalizing LLMs' Assumptions About the User to Calibrate Expectations and Reduce Sycophancy
CHI 2026
·
Myra Cheng
DBLP profile ↗
ORCID search ↗
Whose Knowledge Counts? Co-Designing Community-Centered AI Auditing Tools with Educators in Hawai'i
CHI 2026
·
Dora Zhao
DBLP profile ↗
ORCID search ↗
Aligning Language Models with Demonstrated Feedback
ICLR 2025
·
Omar Shaikh
DBLP profile ↗
ORCID search ↗
Attacking Vision-Language Computer Agents via Pop-ups
ACL 2025
·
Yanzhe Zhang
DBLP profile ↗
ORCID search ↗
Bidirectional Human-AI Alignment: Emerging Challenges and Opportunities
CHI 2025
·
Hua Shen
DBLP profile ↗
ORCID search ↗
Blackbox Model Provenance via Palimpsestic Membership Inference
NeurIPS 2025
·
Rohith Kuditipudi
DBLP profile ↗
ORCID search ↗
Can LLMs Generate Novel Research Ideas? A Large-Scale Human Study with 100+ NLP Researchers
ICLR 2025
·
Chenglei Si
DBLP profile ↗
ORCID search ↗
Creating General User Models from Computer Use
UIST 2025
·
Omar Shaikh
DBLP profile ↗
ORCID search ↗
Culture Cartography: Mapping the Landscape of Cultural Knowledge
EMNLP 2025
·
Caleb Ziems
DBLP profile ↗
ORCID search ↗
Design2Code: Benchmarking Multimodal Code Generation for Automated Front-End Engineering
NAACL 2025
·
Chenglei Si
DBLP profile ↗
ORCID search ↗
Distilling an End-to-End Voice Assistant Without Instruction Training Data
ACL 2025
·
William Barr Held
DBLP profile ↗
ORCID search ↗
EgoNormia: Benchmarking Physical-Social Norm Understanding
ACL 2025
·
MohammadHossein Rezaei
DBLP profile ↗
ORCID search ↗
EquiBench: Benchmarking Large Language Models' Reasoning about Program Semantics via Equivalence Checking
EMNLP 2025
·
Anjiang Wei
DBLP profile ↗
ORCID search ↗
Helping the Helper : Supporting Peer Counselors via AI-Empowered Practice and Feedback
CSCW 2025
·
Shang-Ling Hsu
DBLP profile ↗
ORCID search ↗
Human-AI Collaboration: How AIs Augment Human Teammates
ACL 2025
·
Sherry Wu
DBLP profile ↗
ORCID search ↗
Identifying Unlearned Data in LLMs via Membership Inference Attacks
EMNLP 2025
·
Advit Deepak
DBLP profile ↗
ORCID search ↗
Information Retrieval Induced Safety Degradation in AI Agents
NeurIPS 2025
·
Cheng Yu
DBLP profile ↗
ORCID search ↗
Internal Causal Mechanisms Robustly Predict Language Model Out-of-Distribution Behaviors
ICML 2025
·
Jing Huang
DBLP profile ↗
ORCID search ↗
Knoll: Creating a Knowledge Ecosystem for Large Language Models
UIST 2025
·
Dora Zhao
DBLP profile ↗
ORCID search ↗
Mind the Gap: Static and Interactive Evaluations of Large Audio Models
ACL 2025
·
Minzhi Li
DBLP profile ↗
ORCID search ↗
No Preference Left Behind: Group Distributional Preference Optimization
ICLR 2025
·
Binwei Yao
DBLP profile ↗
ORCID search ↗
OpenCUA: Open Foundations for Computer-Use Agents
NeurIPS 2025
·
Xinyuan Wang
DBLP profile ↗
ORCID search ↗
Position: Towards Bidirectional Human-AI Alignment
NeurIPS 2025
·
Hua Shen
DBLP profile ↗
ORCID search ↗
SPHERE: An Evaluation Card for Human-AI Systems
ACL 2025
·
Dora Zhao
DBLP profile ↗
ORCID search ↗
SWE-bench Multimodal: Do AI Systems Generalize to Visual Software Domains?
ICLR 2025
·
John Yang
DBLP profile ↗
ORCID search ↗
SWE-smith: Scaling Data for Software Engineering Agents
NeurIPS 2025
·
John Yang
DBLP profile ↗
ORCID search ↗
Sketch2Code: Evaluating Vision-Language Models for Interactive Web Design Prototyping
NAACL 2025
·
Ryan Li
DBLP profile ↗
ORCID search ↗
StorySage: Conversational Autobiography Writing Powered by a Multi-Agent Framework
UIST 2025
·
Shayan Talaei
DBLP profile ↗
ORCID search ↗
SynthesizeMe! Inducing Persona-Guided Prompts for Personalized Reward Models in LLMs
ACL 2025
·
Michael J. Ryan
DBLP profile ↗
ORCID search ↗
The Practice of Online Peer Counseling and the Potential for AI-Powered Support Tools
CSCW 2025
·
Tony Wang
DBLP profile ↗
ORCID search ↗
When Models Know More Than They Can Explain: Quantifying Knowledge Transfer in Human-AI Collaboration
NeurIPS 2025
·
Quan Shi
DBLP profile ↗
ORCID search ↗
Are Large Language Models Consistent over Value-laden Questions?
EMNLP 2024
·
Jared Moore
DBLP profile ↗
ORCID search ↗
Benchmarking Machine Translation with Cultural Awareness
EMNLP 2024
·
Binwei Yao
DBLP profile ↗
ORCID search ↗
CultureBank: An Online Community-Driven Knowledge Base Towards Culturally Aware Language Technologies
EMNLP 2024
·
Weiyan Shi
DBLP profile ↗
ORCID search ↗
DARG: Dynamic Evaluation of Large Language Models via Adaptive Reasoning Graph
NeurIPS 2024
·
Zhehao Zhang
DBLP profile ↗
ORCID search ↗
Decoding Susceptibility: Modeling Misbelief to Misinformation Through a Computational Approach
EMNLP 2024
·
Yanchen Liu
DBLP profile ↗
ORCID search ↗
Demystifying Verbatim Memorization in Large Language Models
EMNLP 2024
·
Jing Huang
DBLP profile ↗
ORCID search ↗
DyVal: Dynamic Evaluation of Large Language Models for Reasoning Tasks
ICLR 2024
·
Kaijie Zhu
DBLP profile ↗
ORCID search ↗
Grounding Gaps in Language Model Generations
NAACL 2024
·
Omar Shaikh
DBLP profile ↗
ORCID search ↗
How Johnny Can Persuade LLMs to Jailbreak Them: Rethinking Persuasion to Challenge AI Safety by Humanizing LLMs
ACL 2024
·
Yi Zeng
DBLP profile ↗
ORCID search ↗
Language Agents: Foundations, Prospects, and Risks
EMNLP 2024
·
Yu Su
DBLP profile ↗
ORCID search ↗
MIDDAG: Where Does Our News Go? Investigating Information Diffusion via Community-Level Information Pathways
AAAI 2024
·
Mingyu Derek Ma
DBLP profile ↗
ORCID search ↗
Measuring and Addressing Indexical Bias in Information Retrieval
ACL 2024
·
Caleb Ziems
DBLP profile ↗
ORCID search ↗
Modeling Gender and Dialect Bias in Automatic Speech Recognition
EMNLP 2024
·
Camille Harris
DBLP profile ↗
ORCID search ↗
Multi-Level Feedback Generation with Large Language Models for Empowering Novice Peer Counselors
ACL 2024
·
Alicja Chaszczewicz
DBLP profile ↗
ORCID search ↗
Perceptions of Language Technology Failures from South Asian English Speakers
ACL 2024
·
Faye Holt
DBLP profile ↗
ORCID search ↗
Position: A Safe Harbor for AI Evaluation and Red Teaming
ICML 2024
·
Shayne Longpre
DBLP profile ↗
ORCID search ↗
PrivacyLens: Evaluating Privacy Norm Awareness of Language Models in Action
NeurIPS 2024
·
Yijia Shao
DBLP profile ↗
ORCID search ↗
Rehearsal: Simulating Conflict to Teach Conflict Resolution
CHI 2024
·
Omar Shaikh
DBLP profile ↗
ORCID search ↗
Roleplay-doh: Enabling Domain-Experts to Create LLM-simulated Patients via Eliciting and Adhering to Principles
EMNLP 2024
·
Ryan Louie
DBLP profile ↗
ORCID search ↗
Semi-Truths: A Large-Scale Dataset of AI-Augmented Images for Evaluating Robustness of AI-Generated Image detectors
NeurIPS 2024
·
Anisha Pal
DBLP profile ↗
ORCID search ↗
Silent Signals, Loud Impact: LLMs for Word-Sense Disambiguation of Coded Dog Whistles
ACL 2024
·
Julia Kruk
DBLP profile ↗
ORCID search ↗
Simulated Misinformation Susceptibility (SMISTS): Enhancing Misinformation Research with Large Language Model Simulations
ACL 2024
·
Weicheng Ma
DBLP profile ↗
ORCID search ↗
Social Intelligence Data Infrastructure: Structuring the Present and Navigating the Future
ACL 2024
·
Minzhi Li
DBLP profile ↗
ORCID search ↗
Training Socially Aligned Language Models on Simulated Social Interactions
ICLR 2024
·
Ruibo Liu
DBLP profile ↗
ORCID search ↗
Understanding Online Discussion Across Difference: Insights from Gun Discourse on Reddit
CSCW 2024
·
Rijul Magu
DBLP profile ↗
ORCID search ↗
Unintended Impacts of LLM Alignment on Global Representation
ACL 2024
·
Michael J. Ryan
DBLP profile ↗
ORCID search ↗
What Makes Digital Support Effective? How Therapeutic Skills Affect Clinical Well-Being
CSCW 2024
·
Wenjie Yang
DBLP profile ↗
ORCID search ↗