PPaperPicks

Minlong Peng

10 papers at tracked venues · 7 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. DeCoRL: Decoupling Reasoning Chains via Parallel Sub-Step Generation and Cascaded Reinforcement for Interpretable and Scalable RLHF
  2. Lingua-Graph: A Unified Representation of Cross-Task Common Substructures for Analytic Language Processing
  3. Parameter Importance is Not Static: Evolving Parameter Isolation for Supervised Fine-Tuning
  4. Reason Only When Needed: Efficient Generative Reward Modeling via Model-Internal Uncertainty
  5. Reinforcement Learning Enhanced Muti-hop Reasoning for Temporal Knowledge Question Answering
  6. Why Supervised Fine-Tuning Fails to Learn: A Systematic Study of Incomplete Learning in Large Language Models
  7. Not All Parameters Are Created Equal: Smart Isolation Boosts Fine-Tuning Performance
  8. Structural Reward Model: Enhancing Interpretability, Efficiency, and Scalability in Reward Modeling
  9. Fooling the Textual Fooler via Randomizing Latent Representations
  10. One2Set + Large Language Model: Best Partners for Keyphrase Generation