PPaperPicks

Qiang He

6 papers at tracked venues · 5 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Scaling Behaviors of LLM Reinforcement Learning Post-Training: An Empirical Study in Mathematical Reasoning
  2. DiffGrasp: Whole-Body Grasping Synthesis Guided by Object Motion Using a Diffusion Model
  3. Pareto Multi-objective Alignment for Language Models
    ECML-PKDD 2025 · Qiang He
  4. Adaptive Regularization of Representation Rank as an Implicit Constraint of Bellman Equation
    ICLR 2024 · Qiang He
  5. Advancing DRL Agents in Commercial Fighting Games: Training, Integration, and Agent-Human Alignment
  6. Task Adaptation from Skills: Information Geometry, Disentanglement, and New Objectives for Unsupervised Reinforcement Learning