PPaperPicks

Cheng Qian

University of Illinois Urbana-Champaign, Champaign, IL, USA

26 papers at tracked venues · 20 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Copyright Detective: A Forensic System to Evidence LLMs Flickering Copyright Leakage Risks
  2. CostBench: Evaluating Multi-Turn Cost-Optimal Planning and Adaptation in Dynamic Environments for LLM Tool-Use Agents
  3. Current Agents Fail to Leverage World Model as Tool for Foresight
    ACL 2026 · Cheng Qian
  4. From Word to World: Can Large Language Models be Implicit Text-based World Models?
  5. PEARL: Self-Evolving Assistant for Time Management with Reinforcement Learning
  6. ShortageSim: Simulating Drug Shortages Under Information Asymmetry
  7. TT-SI: Self-Improving LLM Agents with Test-Time Training
  8. Veri-R1: Toward Precise and Faithful Claim Verification via Online Reinforcement Learning
  9. WiNELL: Wikipedia Never-Ending Updating with LLM Agents
  10. DecisionFlow: Advancing Large Language Model as Principled Decision Maker
  11. Distance between Relevant Information Pieces Causes Bias in Long-Context LLMs
  12. EmbodiedBench: Comprehensive Benchmarking Multi-modal Large Language Models for Vision-Driven Embodied Agents
  13. Enhancing Open-Domain Task-Solving Capability of LLMs via Autonomous Tool Integration from GitHub
  14. EscapeBench: Towards Advancing Creative Intelligence of Language Model Agents
    ACL 2025 · Cheng Qian
  15. ISACL: Internal State Analyzer for Copyrighted Training Data Leakage
  16. ModelingAgent: Bridging LLMs and Mathematical Modeling for Real-World Challenges
    EMNLP 2025 · Cheng Qian
  17. MultiAgentBench : Evaluating the Collaboration and Competition of LLM agents
  18. Proactive Agent: Shifting LLM Agents from Reactive Responses to Active Assistance
  19. Rescorla-Wagner Steering of LLMs for Undesired Behaviors over Disproportionate Inappropriate Context
  20. SMART: Self-Aware Agent for Tool Overuse Mitigation
    ACL 2025 · Cheng Qian
  21. SafeSwitch: Steering Unsafe LLM Behavior via Internal Activation Signals
  22. The Law of Knowledge Overshadowing: Towards Understanding, Predicting and Preventing LLM Hallucination
  23. The Right Time Matters: Data Arrangement Affects Zero-Shot Generalization in Instruction Tuning
  24. ToolRL: Reward is All Tool Learning Needs
    NeurIPS 2025 · Cheng Qian
  25. Tell Me More! Towards Implicit User Intention Understanding of Language Model Driven Agents
    ACL 2024 · Cheng Qian
  26. Toolink: Linking Toolkit Creation and Using through Chain-of-Solving on Open-Source Model
    NAACL 2024 · Cheng Qian