PPaperPicks

Weixun Wang

10 papers at tracked venues · 8 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. CE-RM: A Pointwise Generative Reward Model Optimized via Two-Stage Rollout and Unified Criteria
  2. ShopSimulator: Evaluating and Exploring RL-Driven LLM Agent for Shopping Assistants
  3. Think-J: Learning to Think for Generative LLM-as-a-Judge
  4. USB: A Comprehensive and Unified Safety Evaluation Benchmark for Multimodal Large Language Models
  5. 2D-DPO: Scaling Direct Preference Optimization with 2-Dimensional Supervision
  6. Can Large Language Models Detect Errors in Long Chain-of-Thought Reasoning?
  7. Chinese SimpleQA: A Chinese Factuality Evaluation for Large Language Models
  8. OpenRLHF: A Ray-based Easy-to-use, Scalable and High-performance RLHF Framework
  9. ProgCo: Program Helps Self-Correction of Large Language Models
  10. PORTAL: Automatic Curricula Generation for Multiagent Reinforcement Learning