PPaperPicks

Songyang Gao

13 papers at tracked venues · 10 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. AgentGym: Evaluating and Training Large Language Model-based Agents across Diverse Environments
  2. Alleviating Shifted Distribution in Human Preference Alignment through Meta-Learning
  3. Are Your LLMs Capable of Stable Reasoning?
  4. Capability Salience Vector: Fine-grained Alignment of Loss and Capabilities for Downstream Task Scaling Law
  5. CompassVerifier: A Unified and Robust Verifier for LLMs Evaluation and Outcome Reward
  6. Pre-Trained Policy Discriminators are General Reward Models
  7. Semi-off-Policy Reinforcement Learning for Vision-Language Slow-Thinking Reasoning
  8. Inverse-Q*: Token Level Reinforcement Learning for Aligning Large Language Models Without Preference Data
  9. Linear Alignment: A Closed-form Solution for Aligning Human Preferences without Tuning and Feedback
    ICML 2024 · Songyang Gao
  10. LoRAMoE: Alleviating World Knowledge Forgetting in Large Language Models via MoE-Style Plugin
  11. Navigating the OverKill in Large Language Models
  12. RoTBench: A Multi-Level Benchmark for Evaluating the Robustness of Large Language Models in Tool Learning
  13. ToolSword: Unveiling Safety Issues of Large Language Models in Tool Learning Across Three Stages