PPaperPicks

Haonan Li

Mohamed bin Zayed University of Artificial Intelligence, Abu Dhabi, UAE

16 papers at tracked venues · 12 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Control Illusion: The Failure of Instruction Hierarchies in Large Language Models
  2. FinChain: A Symbolic Benchmark for Verifiable Chain-of-Thought Financial Reasoning
  3. Multilingual Idioms in Sentences and Conversations Across High-, Medium-, and Low-Resource Languages
  4. Nanda Family: Open-Weights Generative Large Language Models for Hindi
  5. SCALAR: Scientific Citation-based Live Assessment of Long-context Academic Reasoning
  6. Libra-Leaderboard: Towards Responsible AI through a Balanced Leaderboard of Safety and Capability
    NAACL 2025 · Haonan Li
  7. NAT: Enhancing Agent Tuning with Negative Samples
  8. Revisiting Reinforcement Learning for LLM Reasoning from A Cross-Domain Perspective
  9. ToolGen: Unified Tool Retrieval and Calling via Generation
  10. A Chinese Dataset for Evaluating the Safeguards in Large Language Models
  11. ArabicMMLU: Assessing Massive Multitask Language Understanding in Arabic
  12. CMMLU: Measuring massive multitask language understanding in Chinese
    ACL 2024 · Haonan Li
  13. Demystifying Instruction Mixing for Fine-tuning Large Language Models
  14. EXAMS-V: A Multi-Discipline Multilingual Multimodal Exam Benchmark for Evaluating Vision Language Models
  15. Fact-Checking the Output of Large Language Models via Token-Level Uncertainty Quantification
  16. Web2Code: A Large-scale Webpage-to-Code Dataset and Evaluation Framework for Multimodal LLMs