PPaperPicks

Ming Zhang

15 papers at tracked venues · 11 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Beyond Scaling: Measuring and Predicting the Upper Bound of Knowledge Retention in Language Model Pre-Training
  2. From Scores to Preferences: Redefining Evaluation Paradigm for Speech Quality Reward Modeling
  3. LLMEval-Fair: A Large-Scale Longitudinal Study on Robust and Fair Evaluation of Large Language Models
    ACL 2026 · Ming Zhang
  4. MetaAct-RL: Training Language Models for Reasoning Through Meta-Action-Based Reinforcement Learning
  5. Muse: Towards Reproducible Long-Form Song Generation with Fine-Grained Style Control
  6. Reasoning or Memorization? Unreliable Results of Reinforcement Learning Due to Data Contamination
  7. VRPO: Rethinking Value Modeling for Robust RL under Noisy Supervision in LLM Post-Training
  8. What Makes a Good Speech Tokenizer for LLM-Centric Speech Generation? A Systematic Study
  9. EvaLearn: Quantifying the Learning Capability and Efficiency of LLMs via Sequential Problem Solving
  10. Governance in Motion: Co-evolution of Constitutions and AI models for Scalable Safety
  11. LLMEval-Med: A Real-world Clinical Benchmark for Medical LLMs with Physician Validation
    EMNLP 2025 · Ming Zhang
  12. PFDial: A Structured Dialogue Instruction Fine-tuning Method Based on UML Flowcharts
    ACL 2025 · Ming Zhang
  13. Exploring the Compositional Deficiency of Large Language Models in Mathematical Reasoning Through Trap Problems
  14. LLMEval: A Preliminary Study on How to Evaluate Large Language Models
  15. TransferTOD: A Generalizable Chinese Multi-Domain Task-Oriented Dialogue System with Transfer Capabilities
    EMNLP 2024 · Ming Zhang