PPaperPicks

Linjing Li

14 papers at tracked venues · 7 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Entropy Scheduling in Reinforcement Learning for Large Language Models
  2. Spec-o3: A Tool-Augmented Vision-Language Agent for Rare Celestial Object Candidate Vetting via Automated Spectral Inspection
  3. Beyond the First Error: Process Reward Models for Reflective Mathematical Reasoning
  4. CPE: A New Paradigm for Policy Extraction in Offline Reinforcement Learning
  5. Learning Dynamics in Continual Pre-Training for Large Language Models
  6. Learning Strategy Representation for Imitation Learning in Multi-Agent Games
  7. Learning Theorem Rationale for Improving the Mathematical Reasoning Capability of Large Language Models
  8. Learning When to Think: Shaping Adaptive Reasoning in R1-Style Models via Multi-Stage RL
  9. Offline Meta Reinforcement Learning with Weighted Policy Constraints and Proximal Context Collection
  10. POSITION BIAS MITIGATES POSITION BIAS: Mitigate Position Bias Through Inter-Position Knowledge Distillation
  11. Uncertainty Unveiled: Can Exposure to More In-context Examples Mitigate Uncertainty for Large Language Models?
  12. Unearthing Gems from Stones: Policy Optimization with Negative Sample Augmentation for LLM Reasoning
  13. ELA: Exploited Level Augmentation for Offline Learning in Zero-Sum Games
  14. Unveiling Factual Recall Behaviors of Large Language Models through Knowledge Neurons