PPaperPicks

Qiaozhi He

6 papers at tracked venues · 6 at CORE A* · active 20252026

Venues

Frequent coauthors

Papers

  1. Probing Preference Representations: A Multi-Dimensional Evaluation and Analysis Method for Reward Models
  2. SERM: Self-Evolving Relevance Model with Agent-Driven Learning from Massive Query Streams
  3. SPS: Steering Probability Squeezing for Better Exploration in Reinforcement Learning for Large Language Models
  4. GRAM: A Generative Foundation Reward Model for Reward Generalization
  5. RoVRM: A Robust Visual Reward Model Optimized via Auxiliary Textual Preference Data
  6. StickMotion: Generating 3D Human Motions by Drawing a Stickman