PPaperPicks

Wei Shen

19 papers at tracked venues · 14 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Pre-DPO: Improving Data Utilization in Direct Preference Optimization Using a Guiding Reference Model
  2. SWE-Swiss: A Multi-Task Fine-Tuning and RL Recipe for High-Performance Issue Resolution
  3. A Single-Loop First-Order Algorithm for Linearly Constrained Bilevel Optimization
    NeurIPS 2025 · Wei Shen
  4. Exploring Data Scaling Trends and Effects in Reinforcement Learning from Human Feedback
    NeurIPS 2025 · Wei Shen
  5. Joint Enhancement of Relational Reasoning for Long-Context LLMs
  6. Learning LLM-as-a-Judge for Preference Alignment
  7. Learning from LLM Agents: In-Context Generative Models for Text Casing in E-Commerce Ads
  8. Mitigating Posterior Salience Attenuation in Long-Context LLMs with Positional Contrastive Decoding
  9. On the Training Convergence of Transformers for In-Context Classification of Gaussian Mixtures
    ICML 2025 · Wei Shen
  10. OpenRLHF: A Ray-based Easy-to-use, Scalable and High-performance RLHF Framework
  11. RMB: Comprehensively benchmarking reward models in LLM alignment
  12. Improving Generalization of Alignment with Human Preferences through Group Invariant Learning
  13. Linear Alignment: A Closed-form Solution for Aligning Human Preferences without Tuning and Feedback
  14. LoRAMoE: Alleviating World Knowledge Forgetting in Large Language Models via MoE-Style Plugin
  15. Mitigating Reward Overoptimization via Lightweight Uncertainty Estimation
  16. Reward Modeling Requires Automatic Adjustment Based on Data Quality
  17. Stochastic Smoothed Gradient Descent Ascent for Federated Minimax Optimization
    AISTATS 2024 · Wei Shen
  18. Training Large Language Models for Reasoning through Reverse Curriculum Reinforcement Learning
  19. ViTree: Single-Path Neural Tree for Step-Wise Interpretable Fine-Grained Visual Categorization