PPaperPicks

Liangming Pan

33 papers at tracked venues · 23 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. LEDOM: Reverse Language Model
  2. Towards Intrinsic Interpretability of Large Language Models: A Survey of Design Principles and Architectures
  3. Towards a Mechanistic Understanding of Large Reasoning Models: A Survey of Training, Inference, and Failures
  4. AntiLeakBench: Preventing Data Contamination by Automatically Constructing Benchmarks with Updated Real-World Knowledge
  5. Aristotle: Mastering Logical Reasoning with A Logic-Complete Decompose-Search-Resolve Framework
  6. CausalEval: Towards Better Causal Reasoning in Language Models
  7. Combating Multimodal LLM Hallucination via Bottom-Up Holistic Reasoning
  8. ConciseRL: Conciseness-Guided Reinforcement Learning for Efficient Reasoning Models
  9. Gödel Agent: A Self-Referential Agent Framework for Recursively Self-Improvement
  10. How Is LLM Reasoning Distracted by Irrelevant Context? An Analysis Using a Controlled Benchmark
  11. How do Transformers Learn Implicit Reasoning?
  12. InductionBench: LLMs Fail in the Simplest Complexity Class
  13. Investigating the Transferability of Code Repair for Low-Resource Programming Languages
  14. Long Context vs. RAG: Strategies for Processing Long Documents in LLMs
  15. MuSLR: Multimodal Symbolic Logical Reasoning
  16. RuleArena: A Benchmark for Rule-Guided Reasoning with LLMs in Real-World Scenarios
  17. SeaKR: Self-aware Knowledge Retrieval for Adaptive Retrieval Augmented Generation
  18. TART: An Open-Source Tool-Augmented Framework for Explainable Table-based Reasoning
  19. Towards Temporal-Aware Multi-Modal Retrieval Augemented Generation in Finance
  20. A Survey on Detection of LLMs-Generated Content
  21. AKEW: Assessing Knowledge Editing in the Wild
  22. Factcheck-Bench: Fine-Grained Evaluation Benchmark for Automatic Fact-checkers
  23. Faithful Logical Reasoning via Symbolic Chain-of-Thought
  24. Knowledge of Knowledge: Exploring Known-Unknowns Uncertainty with Large Language Models
  25. MMLONGBENCH-DOC: Benchmarking Long-context Document Understanding with Visualizations
  26. Modeling Dynamic Topics in Chain-Free Fashion by Evolution-Tracking Contrastive Learning and Unassociated Word Exclusion
  27. MultiAgent Collaboration Attack: Investigating Adversarial Attacks in Large Language Model Collaborations via Debate
  28. Position: AI/ML Influencers Have a Place in the Academic Process
  29. Pride and Prejudice: LLM Amplifies Self-Bias in Self-Refinement
  30. SciAgent: Tool-augmented Language Models for Scientific Reasoning
  31. The Knowledge Alignment Problem: Bridging Human and External Knowledge for Large Language Models
  32. Towards Verifiable Generation: A Benchmark for Knowledge-aware Language Model Attribution
  33. Understanding Reasoning Ability of Language Models From the Perspective of Reasoning Paths Aggregation