PPaperPicks

Yutao Zhu

Université de Montréal, Canada

46 papers at tracked venues · 40 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. ATIR: Towards Audio-Text Interleaved Contextual Retrieval
  2. Agentic-R: Learning to Retrieve for Agentic Search
  3. DeepAgent: A General Reasoning Agent with Scalable Toolsets
  4. EnvScaler: Scaling Tool-Interactive Environments for LLM Agent via Programmatic Synthesis
  5. Evaluating the Factuality of Large Language Models Using Multiple Plug-and-Play Fact Sources
  6. FinSight: Towards Real-World Financial Deep Research
  7. HiRA: Decoupling Planning and Execution with Hierarchical Reasoning in Deep Search
  8. Hide and Seek with LLMs: An Adversarial Game for Sneaky Error Generation and Self-Improving Diagnosis
  9. Internalizing Explicit Reasoning into Latent Space for Dense Retrieval
  10. LLM-Generated Text May Harm Your Retrieval! A Robust Detection Strategy for Retrieval-Augmented Generation
  11. Memory Matters More: Event-Centric Memory as a Logic Map for Agent Searching and Reasoning
  12. RLSeek: Evidence-Grounded Reasoning for RAG Hallucination Detection
  13. ReasonRank: Empowering Passage Ranking with Strong Reasoning Ability
  14. R⌃3AG: Retriever Routing for Retrieval-Augmented Generation
  15. Tool-Star: Empowering Multi-Tool Collaborative Web Agent via Reinforcement Learning
  16. Toward Generalized Web Agent Training: A Deep Dive into Entropy-Balanced Reinforcement Learning
  17. Web Sitemap Knowledge Can Enhance Autonomous Browsing
  18. CoRanking: Collaborative Ranking with Small and Large Ranking Agents
  19. Descriptive and Discriminative Document Identifiers for Generative Retrieval
  20. Embedding Prior Task-specific Knowledge into Language Models for Context-aware Document Ranking
  21. Enhancing LLM Text Detection with Retrieved Contexts and Logits Distribution Consistency
  22. FlashRAG: A Modular Toolkit for Efficient Retrieval-Augmented Generation Research
  23. From Pixels to Tokens: Revisiting Object Hallucinations in Large Vision-Language Models
  24. Hierarchical Document Refinement for Long-context Retrieval-augmented Generation
  25. LLMs + Persona-Plug = Personalized LLMs
  26. Little Giants: Synthesizing High-Quality Embedding Data at Scale
  27. One Token Can Help! Learning Scalable and Pluggable Virtual Tokens for Retrieval-Augmented Large Language Models
    AAAI 2025 · Yutao Zhu
  28. Progressive Multimodal Reasoning via Active Retrieval
  29. RAG-Critic: Leveraging Automated Critic-Guided Agentic Workflow for Retrieval Augmented Generation
  30. Search-o1: Agentic Search-Enhanced Large Reasoning Models
  31. Single LLM, Multiple Roles: A Unified Retrieval-Augmented Generation Framework Using Role-Specific Token Optimization
    EMNLP 2025 · Yutao Zhu
  32. Sliding Windows Are Not the End: Exploring Full Ranking with Long-Context Large Language Models
  33. Toward Verifiable Instruction-Following Alignment for Retrieval Augmented Generation
  34. Towards Effective and Efficient Continual Pre-training of Large Language Models
  35. Understand What LLM Needs: Dual Preference Alignment for Retrieval-Augmented Generation
  36. WebThinker: Empowering Large Reasoning Models with Deep Research Capability
  37. YuLan-Mini: Pushing the Limits of Open Data-efficient Language Model
  38. mmE5: Improving Multimodal Multilingual Embeddings via High-quality Synthetic Data
  39. An Integrated Data Processing Framework for Pretraining Foundation Models
  40. BIDER: Bridging Knowledge Inconsistency for Efficient Retrieval-Augmented LLMs via Key Supporting Evidence
  41. CL4DIV: A Contrastive Learning Framework for Search Result Diversification
  42. INTERS: Unlocking the Power of Large Language Models in Search with Instruction Tuning
    ACL 2024 · Yutao Zhu
  43. Information Retrieval Meets Large Language Models
  44. JDivPS: A Diversified Product Search Dataset
  45. Mining Exploratory Queries for Conversational Search
  46. Small Models, Big Insights: Leveraging Slim Proxy Models To Decide When and What to Retrieve for LLMs