PPaperPicks

Yong Jiang

Alibaba Group, DAMO Academy, China

34 papers at tracked venues · 17 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Benchmarking Temporal Reasoning and Alignment Across Chinese Dynasties
  2. BrowseConf: Confidence-Guided Test-Time Scaling for Web Agents
  3. Nested Browser-Use Learning for Agentic Information Seeking
  4. STORM: A Spatio-Temporal Factor Model Based on Dual Vector Quantized Variational Autoencoders for Financial Trading
  5. Towards General Agentic Intelligence via Environment Scaling
  6. WebAnchor: Anchoring Agent Planning to Stabilize Long-Horizon Web Reasoning
  7. Agentic Knowledgeable Self-awareness
  8. Benchmarking Agentic Workflow Generation
  9. Benchmarking Multimodal Retrieval Augmented Generation with Dynamic VQA Dataset and Self-adaptive Planning Agent
  10. DecoupleSearch: Decouple Planning and Search via Hierarchical Reward Modeling
  11. Detecting Knowledge Boundary of Vision Large Language Models by Sampling-Based Inference
  12. EvolveSearch: An Iterative Self-Evolving Search Agent
  13. KBM: Delineating Knowledge Boundary for Adaptive Retrieval in Large Language Models
  14. LaRA: Benchmarking Retrieval-Augmented Generation and Long-Context LLMs - No Silver Bullet for LC or RAG Routing
  15. OmniThink: Expanding Knowledge Boundaries in Machine Writing through Thinking
  16. ProductAgent: Benchmarking Conversational Product Search Agent with Asking Clarification Questions
  17. Supportiveness-based Knowledge Rewriting for Retrieval-augmented Language Modeling
  18. SynWorld: Virtual Scenario Synthesis for Agentic Action Knowledge Refinement
  19. Unfolding the Headline: Iterative Self-Questioning for News Retrieval and Timeline Summarization
  20. WebDancer: Towards Autonomous Information Seeking Agency
  21. WebWalker: Benchmarking LLMs in Web Traversal
  22. Agent Planning with World Knowledge Model
  23. EcomGPT: Instruction-Tuning Large Language Models with Chain-of-Task Tasks for E-commerce
  24. Effective Demonstration Annotation for In-Context Learning via Language Model-Based Determinantal Point Process
  25. Exploring Key Point Analysis with Pairwise Generation and Graph Partitioning
  26. FactCHD: Benchmarking Fact-Conflicting Hallucination Detection
  27. Improving Retrieval Augmented Open-Domain Question-Answering with Vectorized Contexts
  28. Knowledge Mechanisms in Large Language Models: A Survey and Perspective
  29. Query Routing for Homogeneous Tools: An Instantiation in the RAG Scenario
  30. RaFe: Ranking Feedback Improves Query Rewriting for RAG
  31. Retrieved In-Context Principles from Previous Mistakes
  32. SeqGPT: An Out-of-the-Box Large Language Model for Open Domain Sequence Understanding
  33. Three Heads Are Better than One: Improving Cross-Domain NER with Progressive Decomposed Network
  34. WISE: Rethinking the Knowledge Memory for Lifelong Model Editing of Large Language Models