PPaperPicks

Jiafeng Guo

Chinese Academy of Sciences, Institute of Computing Technology, Beijing, China

72 papers at tracked venues · 49 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AdversarialCoT: Single-Document Retrieval Poisoning for LLM Reasoning
  2. An Iterative Utility Judgment Framework Inspired by Philosophical Relevance via LLMs
  3. Beyond Dialogue Time: Temporal Semantic Memory for Personalized LLM Agents
  4. Beyond Relevance: Utility-Centric Retrieval in the LLM Era
  5. Beyond the Crowd: LLM-Augmented Community Notes for Governing Health Misinformation
  6. Compete to Complete: Co-opetition Adversarial Learning for Retrieval-Augmented Generation
  7. Compressing then Matching: An Efficient Pre-training Paradigm for Multimodal Embedding
  8. Detoxification for LLM: From Dataset Itself
  9. Distilling Large Embeddings via Hyperspherical Householder Quantization
  10. Dynamic Prototype-Augmented Stance Detection: Learning from the Seen to Reason about the Unseen
  11. Déjà Vu of Strange Stickers! Enhancing Out-of-Distribution Robustness in Sticker Retrieval via Cross-Modal Intent Alignment
  12. Gated Differentiable Working Memory for Long-Context Language Modeling
  13. How Do LLM-Generated Texts Impact Term-Based Retrieval Models?
  14. Identify-Conceptualize-Align: A Schema-Adaptive Framework for Unified Entity Recognition and Event Detection
  15. Is a Busy Search Agent a Good One? Overthinking and Overretrieval at Scale
  16. Iterative Structured Pruning for Large Language Models with Multi-Domain Calibration
  17. LongRanker: Efficient One-Pass Document Reranking with Long-Context Large Language Models
  18. Lost in Decomposition: Analyzing and Mitigating the Limitations of Long Context Methods via Context Dependency
  19. Modeling Human-Like Cognition for Stance Detection: Integrating Intuitive Judgment and Analytical Reasoning
  20. One-Pass Decoding for Generative Recommendation with WFST-Constrained A* Search
  21. Reconstructing Content with Collaborative Attention for Universal Multimodal Representation Learning
  22. RouteRAG: Efficient Retrieval-Augmented Generation from Text and Graph via Reinforcement Learning
  23. Stop Hardening Everything: A Training-Free Neuron-Level Defense for Neural Ranking Models
  24. Thinking Forward and Backward: Multi-Objective Reinforcement Learning for Retrieval-Augmented Reasoning
  25. Towards Knowledgeable Deep Research: Framework and Benchmark
  26. A Generative Framework for Personalized Sticker Retrieval
  27. A Survey of Link Prediction in N-ary Knowledge Graphs
  28. An Empirical Study of Evaluating Long-form Question Answering
  29. Attack-in-the-Chain: Bootstrapping Large Language Models for Attacks Against Black-Box Neural Ranking Models
  30. Boosting Retrieval-Augmented Generation with Generation-Augmented Retrieval: A Co-Training Approach
  31. Bridging Queries and Tables through Entities in Open-Domain Table Retrieval
  32. CLIPure: Purification in Latent Space via CLIP for Adversarially Robust Zero-Shot Classification
  33. Evaluating Implicit Bias in Large Language Models by Attacking From a Psychometric Perspective
  34. G2S: A General-to-Specific Learning Framework for Temporal Knowledge Graph Forecasting with Large Language Models
  35. Generative Retrieval for Book Search
  36. KnowCoder-X: Boosting Multilingual Information Extraction via Code
  37. MPRF: Interpretable Stance Detection through Multi-Path Reasoning Framework
  38. MVAM: Multi-View Attention Method for Fine-Grained Image-Text Matching
  39. On the Robustness of Generative Information Retrieval Models: An Out-of-Distribution Perspective
  40. On the Scaling of Robustness and Effectiveness in Dense Retrieval
  41. QUITO-X: A New Perspective on Context Compression from the Information Bottleneck Theory
  42. Robust Information Retrieval
  43. Robust-IR @ SIGIR 2025: The First Workshop on Robust Information Retrieval
  44. Tailoring Table Retrieval from a Field-aware Hybrid Matching Perspective
  45. Text2Sql: Pure Fine-Tuning and Pure Knowledge Distillation
  46. The Silent Saboteur: Imperceptible Adversarial Attacks against Black-Box Retrieval-Augmented Generation Systems
  47. Towards Event Extraction with Massive Types: LLM-based Collaborative Annotation and Partitioning Extraction
  48. Towards Fully Exploiting LLM Internal States to Enhance Knowledge Boundary Perception
  49. Towards Robust Universal Information Extraction: Dataset, Evaluation, and Solution
  50. Unbiased Learning to Rank with Query-Level Click Propensity Estimation: Beyond Pointwise Observation and Relevance
  51. Utility-Focused LLM Annotation for Retrieval and Retrieval-Augmented Generation
  52. A Multi-Granularity-Aware Aspect Learning Model for Multi-Aspect Dense Retrieval
  53. Are Large Language Models Good at Utility Judgments?
  54. Bootstrapped Pre-training with Dynamic Identifier Prediction for Generative Retrieval
  55. CausalDiff: Causality-Inspired Disentanglement via Diffusion Model for Adversarial Defense
  56. Classifier Guidance Enhances Diffusion-Based Adversarial Purification by Preserving Predictive Information
  57. Controlling Risk of Retrieval-augmented Generation: A Counterfactual Prompting Framework
  58. Fake News in Sheep's Clothing: Robust Fake News Detection Against LLM-Empowered Style Attacks
  59. GaQR: An Efficient Generation-augmented Question Rewriter
  60. Generative Retrieval Meets Multi-Graded Relevance
  61. KnowCoder: Coding Structured Knowledge into LLMs for Universal Information Extraction
  62. LINKAGE: Listwise Ranking among Varied-Quality References for Non-Factoid QA Evaluation via LLMs
  63. Look Globally and Reason: Two-stage Path Reasoning over Sparse Knowledge Graphs
  64. MORE: Multi-mOdal REtrieval Augmented Generative Commonsense Reasoning
  65. Multi-granular Adversarial Attacks against Black-box Neural Ranking Models
  66. Perturbation-Invariant Adversarial Training for Neural Ranking Models: Improving the Effectiveness-Robustness Trade-Off
  67. Pretraining Data Detection for Large Language Models: A Divergence-based Calibration Method
  68. Recent Advances in Generative Information Retrieval
  69. Recent Advances in Generative Information Retrieval
  70. RoCEL: Advancing Table Entity Linking through Distinctive Row and Column Contexts
  71. Robust Information Retrieval
  72. When Do LLMs Need Retrieval Augmentation? Mitigating LLMs' Overconfidence Helps Retrieval Augmentation