PPaperPicks

Damai Dai

8 papers at tracked venues · 6 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models
  2. Large Language Models Struggle with Unreasonability in Math Problems
  3. Exploring Activation Patterns of Parameters in Language Models
  4. Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention
  5. A Survey on In-context Learning
  6. DeepSeekMoE: Towards Ultimate Expert Specialization in Mixture-of-Experts Language Models
    ACL 2024 · Damai Dai
  7. Let the Expert Stick to His Last: Expert-Specialized Fine-Tuning for Sparse Architectural Large Language Models
  8. Math-Shepherd: Verify and Reinforce LLMs Step-by-step without Human Annotations