PPaperPicks

Pengfei Cao

37 papers at tracked venues · 28 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Bias-Restrained Prefix Representation Finetuning for Mathematical Reasoning
  2. Do LLMs Know Tool Irrelevance? Demystifying Structural Alignment Bias in Tool Invocations
  3. Empowering GUI Agents via Autonomous Experience Exploration and Hindsight Experience Utilization for Task Planning
  4. From Signal Degradation to Computation Collapse: Uncovering the Two Failure Modes of LLM Quantization
  5. Learning How to Remember: A Meta-Cognitive Management Method for Structured and Transferable Agent Memory
  6. Lightweight Haar Wavelet Subband Pruning for LLMs
  7. Look Light, Think Heavy: What Multimodal Chain-of-Thought Reasoning Can and Cannot Do
  8. Task-Stratified Knowledge Scaling Laws for Post-Training Quantized Large Language Models
  9. A Troublemaker with Contagious Jailbreak Makes Chaos in Honest Towns
  10. Agent-RewardBench: Towards a Unified Benchmark for Reward Modeling across Perception, Planning, and Safety in Real-World Multimodal Agents
  11. Beyond Under-Alignment: Atomic Preference Enhanced Factuality Tuning for Large Language Models
  12. CITI: Enhancing Tool Utilizing Ability in Large Language Models Without Sacrificing General Performance
  13. Cracking Factual Knowledge: A Comprehensive Analysis of Degenerate Knowledge Neurons in Large Language Models
  14. DTELS: Towards Dynamic Granularity of Timeline Summarization
  15. Evaluating Personalized Tool-Augmented LLMs from the Perspectives of Personalization and Proactivity
  16. Know-MRI: A Knowledge Mechanisms Revealer&Interpreter for Large Language Models
  17. Knowledge Localization: Mission Not Accomplished? Enter Query Localization!
  18. Knowledge in Superposition: Unveiling the Failures of Lifelong Knowledge Editing for Large Language Models
  19. M2Edit: Locate and Edit Multi-Granularity Knowledge in Multimodal Large Language Model
  20. MIRAGE: Evaluating and Explaining Inductive Reasoning Process in Language Models
  21. RAG-RewardBench: Benchmarking Reward Models in Retrieval Augmented Generation for Preference Alignment
  22. Revealing the Deceptiveness of Knowledge Editing: A Mechanistic Analysis of Superficial Editing
  23. The Knowledge Microscope: Features as Better Analytical Lenses than Neurons
  24. Towards Better Chain-of-Thought: A Reflection on Effectiveness and Faithfulness
  25. Towards Robust Knowledge Unlearning: An Adversarial Framework for Assessing and Improving Unlearning Robustness in Large Language Models
  26. AgentsCourt: Building Judicial Decision-Making Agents with Court Debate Simulation and Legal Knowledge Augmentation
  27. Cutting Off the Head Ends the Conflict: A Mechanism for Interpreting and Mitigating Knowledge Conflicts in Language Models
  28. Focus on Your Question! Interpreting and Mitigating Toxic CoT Problems in Commonsense Reasoning
  29. Journey to the Center of the Knowledge Neurons: Discoveries of Language-Independent Knowledge Neurons and Degenerate Knowledge Neurons
  30. LINKED: Eliciting, Filtering and Integrating Knowledge in Large Language Model for Commonsense Reasoning
  31. MULFE: A Multi-Level Benchmark for Free Text Model Editing
  32. Oasis: Data Curation and Assessment System for Pretraining of Large Language Models
  33. RWKU: Benchmarking Real-World Knowledge Unlearning for Large Language Models
  34. Unlocking the Future: Exploring Look-Ahead Planning Mechanistic Interpretability in Large Language Models
  35. Whispers that Shake Foundations: Analyzing and Mitigating False Premise Hallucinations in Large Language Models
  36. WilKE: Wise-Layer Knowledge Editor for Lifelong Knowledge Editing
  37. ZhuJiu-Knowledge: A Fairer Platform for Evaluating Multiple Knowledge Types in Large Language Models