PPaperPicks

Daxin Jiang

17 papers at tracked venues · 14 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. GitTaskBench: A Benchmark for Code Agents Solving Real-World Tasks Through Code Repository Leveraging
  2. Learning to Compress: Unlocking the Potential of Large Language Models for Text Representation
  3. PRIME: A Process-Outcome Alignment Benchmark for Verifiable Reasoning in Mathematics and Engineering
  4. Beyond the First Error: Process Reward Models for Reflective Mathematical Reasoning
  5. GUI Exploration Lab: Enhancing Screen Navigation in Agents via Multi-Turn Reinforcement Learning
  6. LAVa: Layer-wise KV Cache Eviction with Dynamic Budget Allocation
  7. Open Vision Reasoner: Transferring Linguistic Cognitive Behavior for Visual Reasoning
  8. Open-Reasoner-Zero: An Open Source Approach to Scaling Up Reinforcement Learning on the Base Model
  9. Perception-R1: Pioneering Perception Policy with Reinforcement Learning
  10. Predictable Scale (Part II) - Farseer: A Refined Scaling Law in LLMs
  11. Unearthing Gems from Stones: Policy Optimization with Negative Sample Augmentation for LLM Reasoning
  12. ADAM: Dense Retrieval Distillation with Adaptive Dark Examples
  13. Fine-Grained Distillation for Long Document Retrieval
  14. Retrieval-Augmented Retrieval: Large Language Models are Strong Zero-Shot Retriever
  15. Synergistic Interplay between Search and Large Language Models for Information Retrieval
  16. WizardCoder: Empowering Code Large Language Models with Evol-Instruct
  17. WizardLM: Empowering Large Pre-Trained Language Models to Follow Complex Instructions