PPaperPicks

Minlie Huang

Tsinghua University, Beijing, China

88 papers at tracked venues · 71 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. DPRM: A Dual Implicit Process Reward Model in Multi-Hop Question Answering
  2. Data Efficient RLVR via Off-Policy Influence Guidance
  3. Glyph: Scaling Context Windows via Visual-Text Compression
  4. HoWToBench: Holistic Evaluation for LLM's Capability in Human-level Writing using Tree of Writing
  5. How Should We Enhance the Safety of Large Reasoning Models: An Empirical Study
  6. IF-CRITIC: Towards a Fine-Grained LLM Critic for Instruction-Following Evaluation
  7. IF-RewardBench: Benchmarking Judge Models for Instruction-Following Evaluation
  8. LASA: Language-Agnostic Semantic Alignment at the Semantic Bottleneck for LLM Safety
  9. New Terms, New Toxicity: Consensus-based Chinese Neologism Toxicity Detection via Search-Augmented LLMs
  10. PsychePass: Calibrating LLM Therapeutic Competence via Trajectory-Anchored Tournaments
  11. S⌃4: Operationalizing Speech Act Theory for Strategic Semi-Structured Psychiatric Interview
  12. The Side Effects of Being Smart: Safety Risks in MLLMs' Multi-Image Reasoning
  13. Unveiling the Landscape of Clinical Depression Assessment: From Behavioral Signatures to Psychiatric Reasoning
  14. VisionReward: Fine-Grained Multi-Dimensional Human Preference Learning for Image and Video Generation
  15. WALKSAFE: Risk-aware Graph Random Walk with Bi-GRPO for LLM Safety
  16. When Smiley Turns Hostile: Interpreting How Emojis Trigger LLMs' Toxicity
  17. Ψ-Arena: Interactive Assessment and Optimization of LLM-based Psychological Counselors with Tripartite Feedback
  18. "I've Decided to Leak": Probing Internals Behind Prompt Leakage Intents
  19. A Survey of Post-Training Scaling in Large Language Models
  20. AGD: Adversarial Game Defense Against Jailbreak Attacks in Large Language Models
  21. Advancing Collaborative Debates with Role Differentiation through Multi-Agent Reinforcement Learning
  22. Adversary-Aware DPO: Enhancing Safety Alignment in Vision Language Models via Adversarial Training
  23. Battling against Tough Resister: Strategy Planning with Adversarial Game for Non-collaborative Dialogues
  24. CharacterBench: Benchmarking Character Customization of Large Language Models
  25. CodePlan: Unlocking Reasoning Potential in Large Language Models by Scaling Code-form Planning
  26. Crisp: Cognitive Restructuring of Negative Thoughts through Multi-turn Supportive Dialogues
  27. DCMKC: A Dual Consistency Matching Approach for Multi-hop Question Answering in LLMs
  28. DELMAN: Dynamic Defense Against Large Language Model Jailbreaking with Model Editing
  29. DPGA-TextSyn: Differentially Private Genetic Algorithm for Synthetic Text Generation
  30. DYNTEXT: Semantic-Aware Dynamic Text Sanitization for Privacy-Preserving LLM Inference
  31. Data Selection via Optimal Control for Language Models
  32. DiffusionAttacker: Diffusion-Driven Prompt Manipulation for LLM Jailbreak
  33. Exploring Multimodal Challenges in Toxic Chinese Detection: Taxonomy, Benchmark, and Findings
  34. Guiding not Forcing: Enhancing the Transferability of Jailbreaking Attacks on LLMs via Removing Superfluous Constraints
  35. HCDS: Hierarchical Clustering for Cold-Start Few-Shot Data Selection
  36. HPSS: Heuristic Prompting Strategy Search for LLM Evaluators
  37. Internal Value Alignment in Large Language Models through Controlled Value Vector Activation
  38. JPS: Jailbreak Multimodal Large Language Models with Collaborative Visual Perturbation and Textual Steering
  39. Language Models Learn to Mislead Humans via RLHF
  40. LegalAgentBench: Evaluating LLM Agents in Legal Domain
  41. LogicGame: Benchmarking Rule-Based Reasoning Abilities of Large Language Models
  42. LongSafety: Evaluating Long-Context Safety of Large Language Models
  43. MAGI: Multi-Agent Guided Interview for Psychiatric Assessment
  44. MAPS: Advancing Multi-Modal Reasoning in Expert-Level Physical Science
  45. MHALO: Evaluating MLLMs as Fine-grained Hallucination Detectors
  46. MiniPLM: Knowledge Distillation for Pre-training Language Models
  47. Model Extrapolation Expedites Alignment
  48. Reframe Your Life Story: Interactive Narrative Therapist and Innovative Moment Assessment with Large Language Models
  49. SPaR: Self-Play with Tree-Search Refinement to Improve Instruction-Following in Large Language Models
  50. SS-GEN: A Social Story Generation Framework with Large Language Models
  51. Scenario-independent Uncertainty Estimation for LLM-based Question Answering via Factor Analysis
  52. ShieldVLM: Safeguarding the Multimodal Implicit Toxicity via Deliberative Reasoning with LVLMs: ShieldVLM
  53. SocialEval: Evaluating Social Intelligence of Large Language Models
  54. SocialSim: Towards Socialized Simulation of Emotional Support Conversation
  55. Speculating LLMs' Chinese Training Data Pollution from Their Tokens
  56. Training Language Model to Critique for Better Refinement
  57. Understanding the Dark Side of LLMs' Intrinsic Self-Correction
  58. VPO: Aligning Text-to-Video Generation Models with Prompt Optimization
  59. 360°REA: Towards A Reusable Experience Accumulation with 360° Assessment for Multi-Agent System
  60. AMOR: A Recipe for Building Adaptable Modular Knowledge Agents Through Process Feedback
  61. ASETF: A Novel Method for Jailbreak Attack on LLMs through Translate Suffix Embeddings
  62. AgentBench: Evaluating LLMs as Agents
  63. AlignBench: Benchmarking Chinese Alignment of Large Language Models
  64. AutoDetect: Towards a Unified Framework for Automated Weakness Detection in Large Language Models
  65. Benchmarking Complex Instruction-Following with Multiple Constraints Composition
  66. Black-Box Prompt Optimization: Aligning Large Language Models without Model Training
  67. COKE: A Cognitive Knowledge Graph for Machine Theory of Mind
  68. CharacterGLM: Customizing Social Characters with Large Language Models
  69. CritiqueLLM: Towards an Informative Critique Generation Model for Evaluation of Large Language Model Generation
  70. DC-Instruct: An Effective Framework for Generative Multi-intent Spoken Language Understanding
  71. Defending Large Language Models Against Jailbreaking Attacks Through Goal Prioritization
  72. Depression Detection in Clinical Interviews with LLM-Empowered Structural Element Graph
  73. EmoBench: Evaluating the Emotional Intelligence of Large Language Models
  74. Instruction Pre-Training: Language Models are Supervised Multitask Learners
  75. Language Model Decoding as Direct Metrics Optimization
  76. Language Models Hallucinate, but May Excel at Fact Verification
  77. Large Language Models Are Not Robust Multiple Choice Selectors
  78. Learning Task Decomposition to Assist Humans in Competitive Programming
  79. MiniLLM: Knowledge Distillation of Large Language Models
  80. Mixture-of-Modules: Reinventing Transformers as Dynamic Assemblies of Modules
  81. On Prompt-Driven Safeguarding for Large Language Models
  82. Perception of Knowledge Boundary for Large Language Models through Semi-open-ended Question Answering
  83. SafetyBench: Evaluating the Safety of Large Language Models
  84. ShieldLM: Empowering LLMs as Aligned, Customizable and Explainable Safety Detectors
  85. Thoughts to Target: Enhance Planning for Target-driven Conversation
  86. ToMBench: Benchmarking Theory of Mind in Large Language Models
  87. ToRA: A Tool-Integrated Reasoning Agent for Mathematical Problem Solving
  88. Towards Efficient Exact Optimization of Language Model Alignment