P
PaperPicks
Conferences
Minlie Huang
Tsinghua University, Beijing, China
88 papers at tracked venues · 71 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID 0000-0001-7111-1849 ↗
Google Scholar ↗
Homepage ↗
Venues
ACL
×41
EMNLP
×15
ICLR
×11
AAAI
×9
NeurIPS
×3
ACM MM
×2
ICML
×2
NAACL
×2
ICCV
×1
SIGIR
×1
WWW
×1
Frequent coauthors
Jiale Cheng
DBLP profile ↗
ORCID search ↗
×5
Zhuang Chen
DBLP profile ↗
ORCID search ↗
×5
Zhexin Zhang
DBLP profile ↗
ORCID search ↗
×4
Bosi Wen
DBLP profile ↗
ORCID search ↗
×4
Jinfeng Zhou
DBLP profile ↗
ORCID search ↗
×4
Shiyao Cui
DBLP profile ↗
ORCID search ↗
×3
Jiaxin Wen
DBLP profile ↗
ORCID search ↗
×3
Yuxian Gu
DBLP profile ↗
ORCID search ↗
×3
Chujie Zheng
DBLP profile ↗
ORCID search ↗
×3
Xinyi Wang
DBLP profile ↗
ORCID search ↗
×2
Erle Zhu
DBLP profile ↗
ORCID search ↗
×2
Junxiao Yang
DBLP profile ↗
ORCID search ↗
×2
Papers
DPRM: A Dual Implicit Process Reward Model in Multi-Hop Question Answering
AAAI 2026
·
Xinyi Wang
DBLP profile ↗
ORCID search ↗
Data Efficient RLVR via Off-Policy Influence Guidance
ACL 2026
·
Erle Zhu
DBLP profile ↗
ORCID search ↗
Glyph: Scaling Context Windows via Visual-Text Compression
ACL 2026
·
Jiale Cheng
DBLP profile ↗
ORCID search ↗
HoWToBench: Holistic Evaluation for LLM's Capability in Human-level Writing using Tree of Writing
ACL 2026
·
Andrew Zhuoer Feng
DBLP profile ↗
ORCID search ↗
How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study
ACL 2026
·
Zhexin Zhang
DBLP profile ↗
ORCID search ↗
IF-CRITIC: Towards a Fine-Grained LLM Critic for Instruction-Following Evaluation
ACL 2026
·
Bosi Wen
DBLP profile ↗
ORCID search ↗
IF-RewardBench: Benchmarking Judge Models for Instruction-Following Evaluation
ACL 2026
·
Bosi Wen
DBLP profile ↗
ORCID search ↗
LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety
ACL 2026
·
Junxiao Yang
DBLP profile ↗
ORCID search ↗
New Terms, New Toxicity: Consensus-based Chinese Neologism Toxicity Detection via Search-Augmented LLMs
ACL 2026
·
Shiyao Cui
DBLP profile ↗
ORCID search ↗
PsychePass: Calibrating LLM Therapeutic Competence via Trajectory-Anchored Tournaments
ACL 2026
·
Zhuang Chen
DBLP profile ↗
ORCID search ↗
S⌃4: Operationalizing Speech Act Theory for Strategic Semi-Structured Psychiatric Interview
ACL 2026
·
Guanqun Bi
DBLP profile ↗
ORCID search ↗
The Side Effects of Being Smart: Safety Risks in MLLMs' Multi-Image Reasoning
ACL 2026
·
Renmiao Chen
DBLP profile ↗
ORCID search ↗
Unveiling the Landscape of Clinical Depression Assessment: From Behavioral Signatures to Psychiatric Reasoning
AAAI 2026
·
Zhuang Chen
DBLP profile ↗
ORCID search ↗
VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation
AAAI 2026
·
Jiazheng Xu
DBLP profile ↗
ORCID search ↗
WALKSAFE: Risk-aware Graph Random Walk with Bi-GRPO for LLM Safety
AAAI 2026
·
Shilong Pan
DBLP profile ↗
ORCID search ↗
When Smiley Turns Hostile: Interpreting How Emojis Trigger LLMs' Toxicity
AAAI 2026
·
Shiyao Cui
DBLP profile ↗
ORCID search ↗
Ψ-Arena: Interactive Assessment and Optimization of LLM-based Psychological Counselors with Tripartite Feedback
AAAI 2026
·
Shijing Zhu
DBLP profile ↗
ORCID search ↗
"I've Decided to Leak": Probing Internals Behind Prompt Leakage Intents
EMNLP 2025
·
Jianshuo Dong
DBLP profile ↗
ORCID search ↗
A Survey of Post-Training Scaling in Large Language Models
ACL 2025
·
Hanyu Lai
DBLP profile ↗
ORCID search ↗
AGD: Adversarial Game Defense Against Jailbreak Attacks in Large Language Models
ACL 2025
·
Shilong Pan
DBLP profile ↗
ORCID search ↗
Advancing Collaborative Debates with Role Differentiation through Multi-Agent Reinforcement Learning
ACL 2025
·
Haoran Li
DBLP profile ↗
ORCID search ↗
Adversary-Aware DPO: Enhancing Safety Alignment in Vision Language Models via Adversarial Training
EMNLP 2025
·
Fenghua Weng
DBLP profile ↗
ORCID search ↗
Battling against Tough Resister: Strategy Planning with Adversarial Game for Non-collaborative Dialogues
ACL 2025
·
Haiyang Wang
DBLP profile ↗
ORCID search ↗
CharacterBench: Benchmarking Character Customization of Large Language Models
AAAI 2025
·
Jinfeng Zhou
DBLP profile ↗
ORCID search ↗
CodePlan: Unlocking Reasoning Potential in Large Language Models by Scaling Code-form Planning
ICLR 2025
·
Jiaxin Wen
DBLP profile ↗
ORCID search ↗
Crisp: Cognitive Restructuring of Negative Thoughts through Multi-turn Supportive Dialogues
EMNLP 2025
·
Jinfeng Zhou
DBLP profile ↗
ORCID search ↗
DCMKC: A Dual Consistency Matching Approach for Multi-hop Question Answering in LLMs
EMNLP 2025
·
Xinyi Wang
DBLP profile ↗
ORCID search ↗
DELMAN: Dynamic Defense Against Large Language Model Jailbreaking with Model Editing
ACL 2025
·
Yi Wang
DBLP profile ↗
ORCID search ↗
DPGA-TextSyn: Differentially Private Genetic Algorithm for Synthetic Text Generation
ACL 2025
·
Zhonghao Sun
DBLP profile ↗
ORCID search ↗
DYNTEXT: Semantic-Aware Dynamic Text Sanitization for Privacy-Preserving LLM Inference
ACL 2025
·
Juhua Zhang
DBLP profile ↗
ORCID search ↗
Data Selection via Optimal Control for Language Models
ICLR 2025
·
Yuxian Gu
DBLP profile ↗
ORCID search ↗
DiffusionAttacker: Diffusion-Driven Prompt Manipulation for LLM Jailbreak
EMNLP 2025
·
Hao Wang
DBLP profile ↗
ORCID search ↗
Exploring Multimodal Challenges in Toxic Chinese Detection: Taxonomy, Benchmark, and Findings
ACL 2025
·
Shujian Yang
DBLP profile ↗
ORCID search ↗
Guiding not Forcing: Enhancing the Transferability of Jailbreaking Attacks on LLMs via Removing Superfluous Constraints
ACL 2025
·
Junxiao Yang
DBLP profile ↗
ORCID search ↗
HCDS: Hierarchical Clustering for Cold-Start Few-Shot Data Selection
SIGIR 2025
·
Yuhua Zhao
DBLP profile ↗
ORCID search ↗
HPSS: Heuristic Prompting Strategy Search for LLM Evaluators
ACL 2025
·
Bosi Wen
DBLP profile ↗
ORCID search ↗
Internal Value Alignment in Large Language Models through Controlled Value Vector Activation
ACL 2025
·
Haoran Jin
DBLP profile ↗
ORCID search ↗
JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
ACM MM 2025
·
Renmiao Chen
DBLP profile ↗
ORCID search ↗
Language Models Learn to Mislead Humans via RLHF
ICLR 2025
·
Jiaxin Wen
DBLP profile ↗
ORCID search ↗
LegalAgentBench: Evaluating LLM Agents in Legal Domain
ACL 2025
·
Haitao Li
DBLP profile ↗
ORCID search ↗
LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models
ACL 2025
·
Jiayi Gui
DBLP profile ↗
ORCID search ↗
LongSafety: Evaluating Long-Context Safety of Large Language Models
ACL 2025
·
Yida Lu
DBLP profile ↗
ORCID search ↗
MAGI: Multi-Agent Guided Interview for Psychiatric Assessment
ACL 2025
·
Guanqun Bi
DBLP profile ↗
ORCID search ↗
MAPS: Advancing Multi-Modal Reasoning in Expert-Level Physical Science
ICLR 2025
·
Erle Zhu
DBLP profile ↗
ORCID search ↗
MHALO: Evaluating MLLMs as Fine-grained Hallucination Detectors
ACL 2025
·
Yishuo Cai
DBLP profile ↗
ORCID search ↗
MiniPLM: Knowledge Distillation for Pre-training Language Models
ICLR 2025
·
Yuxian Gu
DBLP profile ↗
ORCID search ↗
Model Extrapolation Expedites Alignment
ACL 2025
·
Chujie Zheng
DBLP profile ↗
ORCID search ↗
Reframe Your Life Story: Interactive Narrative Therapist and Innovative Moment Assessment with Large Language Models
EMNLP 2025
·
Yi Feng
DBLP profile ↗
ORCID search ↗
SPaR: Self-Play with Tree-Search Refinement to Improve Instruction-Following in Large Language Models
ICLR 2025
·
Jiale Cheng
DBLP profile ↗
ORCID search ↗
SS-GEN: A Social Story Generation Framework with Large Language Models
AAAI 2025
·
Yi Feng
DBLP profile ↗
ORCID search ↗
Scenario-independent Uncertainty Estimation for LLM-based Question Answering via Factor Analysis
WWW 2025
·
Zhihua Wen
DBLP profile ↗
ORCID search ↗
ShieldVLM: Safeguarding the Multimodal Implicit Toxicity via Deliberative Reasoning with LVLMs: ShieldVLM
ACM MM 2025
·
Shiyao Cui
DBLP profile ↗
ORCID search ↗
SocialEval: Evaluating Social Intelligence of Large Language Models
ACL 2025
·
Jinfeng Zhou
DBLP profile ↗
ORCID search ↗
SocialSim: Towards Socialized Simulation of Emotional Support Conversation
AAAI 2025
·
Zhuang Chen
DBLP profile ↗
ORCID search ↗
Speculating LLMs' Chinese Training Data Pollution from Their Tokens
EMNLP 2025
·
Qingjie Zhang
DBLP profile ↗
ORCID search ↗
Training Language Model to Critique for Better Refinement
ACL 2025
·
Tianshu Yu
DBLP profile ↗
ORCID search ↗
Understanding the Dark Side of LLMs' Intrinsic Self-Correction
ACL 2025
·
Qingjie Zhang
DBLP profile ↗
ORCID search ↗
VPO: Aligning Text-to-Video Generation Models with Prompt Optimization
ICCV 2025
·
Jiale Cheng
DBLP profile ↗
ORCID search ↗
360°REA: Towards A Reusable Experience Accumulation with 360° Assessment for Multi-Agent System
ACL 2024
·
Shen Gao
DBLP profile ↗
ORCID search ↗
AMOR: A Recipe for Building Adaptable Modular Knowledge Agents Through Process Feedback
NeurIPS 2024
·
Jian Guan
DBLP profile ↗
ORCID search ↗
ASETF: A Novel Method for Jailbreak Attack on LLMs through Translate Suffix Embeddings
EMNLP 2024
·
Hao Wang
DBLP profile ↗
ORCID search ↗
AgentBench: Evaluating LLMs as Agents
ICLR 2024
·
Xiao Liu
DBLP profile ↗
ORCID search ↗
AlignBench: Benchmarking Chinese Alignment of Large Language Models
ACL 2024
·
Xiao Liu
DBLP profile ↗
ORCID search ↗
AutoDetect: Towards a Unified Framework for Automated Weakness Detection in Large Language Models
EMNLP 2024
·
Jiale Cheng
DBLP profile ↗
ORCID search ↗
Benchmarking Complex Instruction-Following with Multiple Constraints Composition
NeurIPS 2024
·
Bosi Wen
DBLP profile ↗
ORCID search ↗
Black-Box Prompt Optimization: Aligning Large Language Models without Model Training
ACL 2024
·
Jiale Cheng
DBLP profile ↗
ORCID search ↗
COKE: A Cognitive Knowledge Graph for Machine Theory of Mind
ACL 2024
·
Jincenzi Wu
DBLP profile ↗
ORCID search ↗
CharacterGLM: Customizing Social Characters with Large Language Models
EMNLP 2024
·
Jinfeng Zhou
DBLP profile ↗
ORCID search ↗
CritiqueLLM: Towards an Informative Critique Generation Model for Evaluation of Large Language Model Generation
ACL 2024
·
Pei Ke
DBLP profile ↗
ORCID search ↗
DC-Instruct: An Effective Framework for Generative Multi-intent Spoken Language Understanding
EMNLP 2024
·
Bowen Xing
DBLP profile ↗
ORCID search ↗
Defending Large Language Models Against Jailbreaking Attacks Through Goal Prioritization
ACL 2024
·
Zhexin Zhang
DBLP profile ↗
ORCID search ↗
Depression Detection in Clinical Interviews with LLM-Empowered Structural Element Graph
NAACL 2024
·
Zhuang Chen
DBLP profile ↗
ORCID search ↗
EmoBench: Evaluating the Emotional Intelligence of Large Language Models
ACL 2024
·
Sahand Sabour
DBLP profile ↗
ORCID search ↗
Instruction Pre-Training: Language Models are Supervised Multitask Learners
EMNLP 2024
·
Daixuan Cheng
DBLP profile ↗
ORCID search ↗
Language Model Decoding as Direct Metrics Optimization
ICLR 2024
·
Haozhe Ji
DBLP profile ↗
ORCID search ↗
Language Models Hallucinate, but May Excel at Fact Verification
NAACL 2024
·
Jian Guan
DBLP profile ↗
ORCID search ↗
Large Language Models Are Not Robust Multiple Choice Selectors
ICLR 2024
·
Chujie Zheng
DBLP profile ↗
ORCID search ↗
Learning Task Decomposition to Assist Humans in Competitive Programming
ACL 2024
·
Jiaxin Wen
DBLP profile ↗
ORCID search ↗
MiniLLM: Knowledge Distillation of Large Language Models
ICLR 2024
·
Yuxian Gu
DBLP profile ↗
ORCID search ↗
Mixture-of-Modules: Reinventing Transformers as Dynamic Assemblies of Modules
EMNLP 2024
·
Zhuocheng Gong
DBLP profile ↗
ORCID search ↗
On Prompt-Driven Safeguarding for Large Language Models
ICML 2024
·
Chujie Zheng
DBLP profile ↗
ORCID search ↗
Perception of Knowledge Boundary for Large Language Models through Semi-open-ended Question Answering
NeurIPS 2024
·
Zhihua Wen
DBLP profile ↗
ORCID search ↗
SafetyBench: Evaluating the Safety of Large Language Models
ACL 2024
·
Zhexin Zhang
DBLP profile ↗
ORCID search ↗
ShieldLM: Empowering LLMs as Aligned, Customizable and Explainable Safety Detectors
EMNLP 2024
·
Zhexin Zhang
DBLP profile ↗
ORCID search ↗
Thoughts to Target: Enhance Planning for Target-driven Conversation
EMNLP 2024
·
Zhonghua Zheng
DBLP profile ↗
ORCID search ↗
ToMBench: Benchmarking Theory of Mind in Large Language Models
ACL 2024
·
Zhuang Chen
DBLP profile ↗
ORCID search ↗
ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving
ICLR 2024
·
Zhibin Gou
DBLP profile ↗
ORCID search ↗
Towards Efficient Exact Optimization of Language Model Alignment
ICML 2024
·
Haozhe Ji
DBLP profile ↗
ORCID search ↗