PPaperPicks

Enyu Zhou

8 papers at tracked venues · 8 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. MM-Doc-R1: Training Agents for Long Document Visual Question Answering through Multi-turn Reinforcement Learning
  2. VRPO: Rethinking Value Modeling for Robust RL under Noisy Supervision in LLM Post-Training
  3. Alleviating Shifted Distribution in Human Preference Alignment through Meta-Learning
  4. Pre-Trained Policy Discriminators are General Reward Models
  5. RMB: Comprehensively benchmarking reward models in LLM alignment
    ICLR 2025 · Enyu Zhou
  6. LCGen: Mining in Low-Certainty Generation for View-consistent Text-to-3D
  7. LoRAMoE: Alleviating World Knowledge Forgetting in Large Language Models via MoE-Style Plugin
  8. StepCoder: Improving Code Generation with Reinforcement Learning from Compiler Feedback