PPaperPicks

Shenzhi Wang

11 papers at tracked venues · 9 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. COIG-P: A High-Quality and Large-Scale Chinese Preference Dataset for Alignment with Human Values
  2. Outcome Accuracy is Not Enough: Aligning the Reasoning Process of Reward Models
  3. Absolute Zero: Reinforced Self-play Reasoning with Zero Data
  4. Beyond the 80/20 Rule: High-Entropy Minority Tokens Drive Effective Reinforcement Learning for LLM Reasoning
    NeurIPS 2025 · Shenzhi Wang
  5. DiveR-CT: Diversity-enhanced Red Teaming Large Language Model Assistants with Relaxing Constraints
  6. Model Surgery: Modulating LLM's Behavior Via Simple Parameter Editing
  7. OS Agents: A Survey on MLLM-based Agents for Computer, Phone and Browser Use
  8. PopAlign: Diversifying Contrasting Patterns for a More Comprehensive Alignment
  9. Boosting LLM Agents with Recursive Contemplation for Effective Deception Handling
    ACL 2024 · Shenzhi Wang
  10. DeeR-VLA: Dynamic Inference of Multimodal Large Language Models for Efficient Robot Execution
  11. PsychoGAT: A Novel Psychological Measurement Paradigm through Interactive Fiction Games with LLM Agents