PPaperPicks

Tian Liang

19 papers at tracked venues · 16 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. CARL: Constraint-Aware Reinforcement Learning for Planning with LLMs
  2. STAPO: Selective Trajectory-Aware Policy Optimization for LLM Agent Training
  3. UrbanGeoEval: A City-Scale Benchmark for Evaluating Large Language Models in Geospatial Reasoning
  4. Competing Large Language Models in Multi-Agent Gaming Environments
  5. Confidence Calibration for Multimodal LLMs: An Empirical Study Through Medical VQA
  6. Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
  7. Do NOT Think That Much for 2+3=? On the Overthinking of Long Reasoning Models
  8. Draft Model Knows When to Stop: Self-Verification Speculative Decoding for Long-Form Generation
  9. MoLE: Decoding by Mixture of Layer Experts Alleviates Hallucination in Large Vision-Language Models
    AAAI 2025 · Tian Liang
  10. RaSA: Rank-Sharing Low-Rank Adaptation
  11. Refuse Whenever You Feel Unsafe: Improving Safety in LLMs via Decoupled Refusal Training
  12. The First Few Tokens Are All You Need: An Efficient and Effective Unsupervised Prefix Fine-Tuning Method for Reasoning Models
  13. Thoughts Are All Over the Place: On the Underthinking of Long Reasoning Models
  14. Trust, But Verify: A Self-Verification Approach to Reinforcement Learning with Verifiable Rewards
  15. Two Experts Are All You Need for Steering Thinking: Reinforcing Cognitive Effort in MoE Reasoning Models Without Additional Training
  16. Addressing Entity Translation Problem via Translation Difficulty and Context Diversity
    ACL 2024 · Tian Liang
  17. CriticBench: Benchmarking LLMs for Critique-Correct Reasoning
  18. Encouraging Divergent Thinking in Large Language Models through Multi-Agent Debate
    EMNLP 2024 · Tian Liang
  19. Querying as Prompt: Parameter-Efficient Learning for Multimodal Language Model
    CVPR 2024 · Tian Liang