PPaperPicks

Juanzi Li

Tsinghua University, Beijing, China

59 papers at tracked venues · 42 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Boundary-Guided Policy Optimization for Memory-efficient RL of Diffusion Large Language Models
  2. Can Large Language Models Effectively Support Decision-Making in Sudden Emergencies?
  3. Chaining the Evidence: Robust Reinforcement Learning for Deep Search Agents with Citation-Aware Rubric Rewards
  4. DeepPrune: Parallel Scaling without Inter-trace Redundancy
  5. From Knowing to Teaching: Scaffolding Pedagogical Decisions for LLM Agent
  6. Personalized Learning Path Planning through Goal-Driven Learner State Modeling
  7. RPC-Bench: A Fine-grained Benchmark for Research Paper Comprehension
  8. SimPBL: A Multi-Agent Framework for Project-Based Learning
  9. SuperWriter: Reflection-Driven Long-Form Generation with Large Language Models
  10. WildReward: Learning Reward Models from In-the-Wild Human Interactions
  11. AGENTIF: Benchmarking Large Language Models Instruction Following Ability in Agentic Scenarios
  12. AdaptThink: Reasoning Models Can Learn When to Think
  13. Agentic Reward Modeling: Integrating Human Preferences with Verifiable Correctness Signals for Reliable Reward Systems
  14. AtomR: Atomic Operator-Empowered Large Language Models for Heterogeneous Knowledge Reasoning
  15. Awaking the Slides: A Tuning-free and Knowledge-regulated AI Tutoring System via Language Model Coordination
  16. CogCoM: A Visual Language Model with Chain-of-Manipulations Reasoning
  17. Constraint Back-translation Improves Complex Instruction Following of Large Language Models
  18. EduCraft: A System for Generating Pedagogical Lecture Scripts from Long-Context Multimodal Presentations
  19. Establishing Trustworthy LLM Evaluation via Shortcut Neuron Analysis
  20. EventSum: A Large-Scale Event-Centric Summarization Dataset for Chinese Multi-News Documents
  21. How do Transformers Learn Implicit Reasoning?
  22. Knowledge-to-Jailbreak: Investigating Knowledge-driven Jailbreaking Attacks for Large Language Models
  23. LLMAEL: Large Language Models are Good Context Augmenters for Entity Linking
  24. LinguaLens: Towards Interpreting Linguistic Mechanisms of Large Language Models via Sparse Auto-Encoder
  25. LongBench v2: Towards Deeper Understanding and Reasoning on Realistic Long-context Multitasks
  26. LongCite: Enabling LLMs to Generate Fine-grained Citations in Long-Context QA
  27. LongReward: Improving Long-context Large Language Models with AI Feedback
  28. LongWriter-V: Enabling Ultra-Long and High-Fidelity Generation in Vision-Language Models
  29. LongWriter: Unleashing 10, 000+ Word Generation from Long Context LLMs
  30. Pre-training Distillation for Large Language Models: A Design Space Exploration
  31. Precise Localization of Memories: A Fine-grained Neuron-level Knowledge Editing Technique for LLMs
  32. RM-Bench: Benchmarking Reward Models of Language Models with Subtlety and Style
  33. SeaKR: Self-aware Knowledge Retrieval for Adaptive Retrieval Augmented Generation
  34. Simulating Classroom Education with LLM-Empowered Agents
  35. SoAy: A Solution-based LLM API-using Methodology for Academic Information Seeking
  36. StoryWriter: A Multi-Agent Framework for Long Story Generation
  37. T1: Advancing Language Model Reasoning through Reinforcement Learning and Inference Scaling
  38. TableLLM: Enabling Tabular Data Manipulation by LLMs in Real Office Usage Scenarios
  39. Towards Understanding Safety Alignment: A Mechanistic Perspective from Safety Neurons
  40. VerIF: Verification Engineering for Reinforcement Learning in Instruction Following
  41. VocQuiz: Vocabulary Question Generation for English Language Education
  42. ADELIE: Aligning Large Language Models on Information Extraction
  43. CogVLM: Visual Expert for Pretrained Language Models
  44. DiaKoP: Dialogue-based Knowledge-oriented Programming for Neural-symbolic Knowledge Base Question Answering
  45. DocEE-zh: A Fine-grained Benchmark for Chinese Document-level Event Extraction
  46. How Proficient Are Large Language Models in Formal Languages? An In-Depth Insight for Knowledge Base Question Answering
  47. KB-Plugin: A Plug-and-play Framework for Large Language Models to Induce Programs over Low-resourced Knowledge Bases
  48. KoLA: Carefully Benchmarking World Knowledge of Large Language Models
  49. LC4EE: LLMs as Good Corrector for Event Extraction
  50. LM-Interview: An Easy-to-use Smart Interviewer System via Knowledge-guided Language Model Exploitation
  51. Let Me Show You Step by Step: An Interpretable Graph Routing Network for Knowledge-based Visual Question Answering
  52. LongAlign: A Recipe for Long Context Alignment of Large Language Models
  53. LongBench: A Bilingual, Multitask Benchmark for Long Context Understanding
  54. MAVEN-ARG: Completing the Puzzle of All-in-One Event Understanding Dataset with Event Argument Annotation
  55. MAVEN-FACT: A Large-scale Event Factuality Detection Dataset
  56. MM-MATH: Advancing Multimodal Math Evaluation with Process Evaluation and Fine-grained Classification
  57. R-Eval: A Unified Toolkit for Evaluating Domain Knowledge of Retrieval Augmented Large Language Models
  58. Transferable and Efficient Non-Factual Content Detection via Probe Training with Offline Consistency Checking
  59. WaterBench: Towards Holistic Evaluation of Watermarks for Large Language Models