PPaperPicks

Zhen Tan

Arizona State University, USA

37 papers at tracked venues · 20 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Explaining the 'Unexplainable' Large Language Models
    WSDM 2026 · Zhen Tan
  2. Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens
  3. Metacognitive Self-Correction for Multi-Agent System via Prototype-Guided Next-Execution Reconstruction
  4. Model Editing as a Double-Edged Sword: Steering Agent Behavior Toward Beneficence or Harm
  5. OR-R1: Automating Modeling and Solving of Operations Research Optimization Problem via Test-Time Reinforcement Learning
  6. The Adaptive Interrogator: Detecting Trojan LLMs in Multi-Agent Systems via Evolved Conversational Strategies
    ACL 2026 ·
    Rana Muhammad Shahroz Khan
  7. ToolPRMBench: Evaluating and Advancing Process Reward Models for Tool-using Agents
  8. <tt>BetaConform</tt>: Efficient MAP Estimation of LLM Ensemble Judgment Performance with Prior Transfer
  9. Agents Under Siege: Breaking Pragmatic Multi-Agent LLM Systems with Optimized Prompt Attacks
  10. AnyMAC: Cascading Flexible Multi-Agent Collaboration via Next-Agent Prediction
  11. Bit-Flip Error Resilience in LLMs: A Comprehensive Analysis and Defense Framework
  12. BrainMAP: Learning Multiple Activation Pathways in Brain Networks
  13. Building Safer Sites: A Large-Scale Multi-Level Dataset for Construction Safety Benchmark
  14. CEB: Compositional Evaluation Benchmark for Fairness in Large Language Models
  15. EQA-RM: A Generative Embodied Reward Model with Test-time Scaling
  16. From Generation to Judgment: Opportunities and Challenges of LLM-as-a-judge
  17. GraphRCG: Self-Conditioned Graph Generation
  18. In Prospect and Retrospect: Reflective Memory Management for Long-term Personalized Dialogue Agents
    ACL 2025 · Zhen Tan
  19. IndustryEQA: Pushing the Frontiers of Embodied Question Answering in Industrial Scenarios
  20. Interpreting Pretrained Language Models via Concept Bottlenecks (Extended Abstract)
    IJCAI 2025 · Zhen Tan
  21. Learning from Diverse Reasoning Paths with Routing and Collaboration
  22. MAPLE: Many-Shot Adaptive Pseudo-Labeling for In-Context Learning
  23. MerRec: A Large-scale Multipurpose Mercari Dataset for Consumer-to-Consumer Recommendation Systems
  24. Multi-Agent Debate for LLM Judges with Adaptive Stability Detection
  25. SCALE: Towards Collaborative Content Analysis in Social Science with Large Language Model Agents and Human Intervention
  26. Separate the Wheat from the Chaff: Winnowing Down Divergent Views in Retrieval Augmented Generation
  27. Task-Aware Resolution Optimization for Visual Large Language Models
  28. Tuning-Free Accountable Intervention for LLM Deployment - a Metacognitive Approach
    AAAI 2025 · Zhen Tan
  29. Window Token Concatenation for Efficient Visual Large Language Models
  30. DALK: Dynamic Co-Augmentation of LLMs and KG to answer Alzheimer's Disease Questions with Scientific Literature
  31. Disinformation Detection: An Evolving Challenge in the Age of LLMs
  32. Facial Affective Behavior Analysis with Instruction Tuning
  33. Glue pizza and eat rocks - Exploiting Vulnerabilities in Retrieval-Augmented Generative Models
    EMNLP 2024 · Zhen Tan
  34. Label Distribution Learning-Enhanced Dual-KNN for Text Classification
  35. Large Language Models for Data Annotation and Synthesis: A Survey
    EMNLP 2024 · Zhen Tan
  36. Sparsity-Guided Holistic Explanation for LLMs with Interpretable Inference-Time Intervention
    AAAI 2024 · Zhen Tan
  37. Thought Graph: Generating Thought Process for Biological Reasoning