PPaperPicks

Muxi Diao

8 papers at tracked venues · 7 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. MedReasoner: Reinforcement Learning Drives Reasoning Grounding from Clinical Thought to Pixel-Level Precision
  2. MemCoRL: Alternating Co-Optimization of Memory Retrieval and Utilization via Collaborative Reinforcement Learning
  3. CS-Bench: A Comprehensive Benchmark for Large Language Models towards Computer Science Mastery
  4. CineTechBench: A Benchmark for Cinematographic Technique Understanding and Generation
  5. SEAS: Self-Evolving Adversarial Safety Optimization for Large Language Models
    AAAI 2025 · Muxi Diao
  6. We-Math: Does Your Large Multimodal Model Achieve Human-like Mathematical Reasoning?
  7. DolphCoder: Echo-Locating Code Large Language Models with Diverse and Multi-Objective Instruction Tuning
  8. How Do Your Code LLMs perform? Empowering Code Instruction Tuning with Really Good Data