PPaperPicks

Le Sun

Chinese Academy of Sciences, Institute of Software, Beijing, China

65 papers at tracked venues · 54 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AI-Salesman: Towards Reliable Large Language Model Driven Telemarketing
  2. Across Programming Language Silos: A Study on Cross-Lingual Retrieval-Augmented Code Generation
  3. All Languages Matter: Understanding and Mitigating Language Bias in Multilingual RAG
  4. Answer First, Evidence Second? Uncovering Hidden Risks in Well-Structured AI Search Summaries
  5. CoCoNUTS: Concentrating on Content while Neglecting Uninformative Textual Styles for AI-Generated Peer Review Detection
  6. CodeRise: Bootstrapping LLMs for Ultra Low-Resource Programming Languages via Progressive Self-Refinement Curriculum
  7. DeepPresenter: Environment-Grounded Reflection for Agentic Presentation Generation
  8. Does Question Really Matter? The Attribution of Answer Bias in LLM Evaluation
  9. Expanding the Boundaries of Vision Prior Knowledge in Multi-modal Large Language Models
  10. Knowing When to Quit: Diagnosing and Training LLMs to Abort Futile Reasoning
  11. MemSearcher: Iterative Memory Integration for Search Agent via End-to-End Reinforcement Learning
  12. Navigating the Infinite Dynamic Web Space: Effective In-Context Exploration via Cognitive Multi-Agent Collaboration
  13. On the Editability of Delta Parameters in Post-Trained Models
  14. PaperRegister: Boosting Flexible-grained Paper Search via Hierarchical Register Indexing
  15. SURE or Not? Investigating Semantic Understanding in Dense Retrieval Models
  16. ScaleBox: Enabling High-Fidelity and Scalable Code Verification for Large Language Models
  17. When Models Outthink Their Safety: Unveiling and Mitigating Self-Jailbreak in Large Reasoning Models
  18. ARise: Towards Knowledge-Augmented Reasoning via Risk-Adaptive Search
  19. AutoAlign: Get Your LLM Aligned with Minimal Annotations
  20. CRUXEVAL-X: A Benchmark for Multilingual Code Reasoning, Understanding and Execution
  21. Cheems: A Practical Guidance for Building and Evaluating Chinese Reward Models from Scratch
  22. Code-SPA: Style Preference Alignment to Large Language Models for Effective and Robust Code Debugging
  23. ConsistentChat: Building Skeleton-Guided Consistent Multi-Turn Dialogues for Large Language Models from Scratch
  24. Critic-CoT: Boosting the Reasoning Abilities of Large Language Model via Chain-of-Thought Critic
  25. DBCopilot: Natural Language Querying over Massive Databases via Schema Routing
  26. DOMAINEVAL: An Auto-Constructed Benchmark for Multi-Domain Code Generation
  27. DeepSolution: Boosting Complex Engineering Solution Design via Tree-based Exploration and Bi-point Thinking
  28. DiffLM: Controllable Synthetic Data Generation via Diffusion Language Models
  29. From Informal to Formal - Incorporating and Evaluating LLMs on Natural Language Requirements to Verifiable Formal Proofs
  30. Large Language Models Often Say One Thing and Do Another
  31. Memorizing is Not Enough: Deep Knowledge Injection Through Reasoning
  32. Multi-Agent Proactive Information Seeking with Adaptive LLM Orchestration for Non-Factoid Question Answering
  33. Not All Terms Matter: Recall-Oriented Adaptive Learning for PLM-aided Query Expansion in Open-Domain Question Answering
  34. On-Policy Self-Alignment with Fine-grained Knowledge Feedback for Hallucination Mitigation
  35. PPTAgent: Generating and Evaluating Presentations Beyond Text-to-Slides
  36. READoc: A Unified Benchmark for Realistic Document Structured Extraction
  37. RMTBench: Benchmarking LLMs Through Multi-Turn User-Centric Role-Playing
  38. Rethinking Reward Model Evaluation: Are We Barking up the Wrong Tree?
  39. Self-Steering Optimization: Autonomous Preference Optimization for Large Language Models
  40. ShortV: Efficient Multimodal Large Language Models by Freezing Visual Tokens in Ineffective Layers
  41. Sparse Latents Steer Retrieval-Augmented Generation
  42. StructRAG: Boosting Knowledge Intensive Reasoning of LLMs via Inference-time Hybrid Information Structurization
  43. Teach Small Models to Reason by Curriculum Distillation
  44. The Devil Is in the Details: Tackling Unimodal Spurious Correlations for Generalizable Multimodal Reward Models
  45. The Linguistic Connectivities Within Large Language Models
  46. The Rise and Down of Babel Tower: Investigating the Evolution Process of Multilingual Code Large Language Model
  47. Transferable Post-training via Inverse Value Learning
  48. Analyze, Generate and Refine: Query Expansion with LLMs for Zero-Shot Open-Domain QA
  49. Benchmarking Large Language Models in Retrieval-Augmented Generation
  50. Chain-of-Rewrite: Aligning Question and Documents for Open-Domain Question Answering
  51. Debiasing In-Context Learning by Instructing LLMs How to Follow Demonstrations
  52. Learning or Self-aligning? Rethinking Instruction Fine-tuning
  53. Mitigating Large Language Model Hallucinations via Autonomous Knowledge Graph-Based Retrofitting
  54. Navigating the Shadows: Unveiling Effective Disturbances for Modern AI Content Detectors
  55. Not All Contexts Are Equal: Teaching LLMs Credibility-aware Generation
  56. Open Grounded Planning: Challenges and Benchmark Construction
  57. PRP-Graph: Pairwise Ranking Prompting to LLMs with Graph Aggregation for Effective Text Re-ranking
  58. REInstruct: Building Instruction Data from Unlabeled Corpus
  59. Rule or Story, Which is a Better Commonsense Expression for Talking with Large Language Models?
  60. Seg2Act: Global Context-aware Action Generation for Document Logical Structuring
  61. Self-Retrieval: End-to-End Information Retrieval with One Large Language Model
  62. SoFA: Shielded On-the-fly Alignment via Priority Rule Following
  63. Spiral of Silence: How is Large Language Model Killing Information Retrieval? - A Case Study on Open Domain Question Answering
  64. StructEval: Deepen and Broaden Large Language Model Assessment via Structured Evaluation
  65. XMC-Agent : Dynamic Navigation over Scalable Hierarchical Index for Incremental Extreme Multi-label Classification