PPaperPicks

Xueqi Cheng

Chinese Academy of Sciences, Institute of Computing Technology, State Key Lab of AI Safety, Beijing, China

126 papers at tracked venues · 81 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AdversarialCoT: Single-Document Retrieval Poisoning for LLM Reasoning
  2. An Iterative Utility Judgment Framework Inspired by Philosophical Relevance via LLMs
  3. AsarRec: Adaptive Sequential Augmentation for Robust Self-supervised Sequential Recommendation
  4. BaseCal: Unsupervised Confidence Calibration via Base Model Signals
  5. Beyond Dialogue Time: Temporal Semantic Memory for Personalized LLM Agents
  6. Compete to Complete: Co-opetition Adversarial Learning for Retrieval-Augmented Generation
  7. D-Models and E-Models: Diversity-Stability Trade-offs in the Sampling Behavior of Large Language Models
  8. Detoxification for LLM: From Dataset Itself
  9. Distilling Large Embeddings via Hyperspherical Householder Quantization
  10. Do We Always Need Query-Level Workflows? Rethinking Agentic Workflow Generation for Multi-Agent Systems
  11. Event-Aware Video Corpus Moment Retrieval
  12. Fairness at a Glance: Can We Audit Model Fairness Before Training Completes?
  13. Gated Differentiable Working Memory for Long-Context Language Modeling
  14. HiddenGuard: Fine-Grained Safe Generation with Specialized Representation Router
  15. How Do LLM-Generated Texts Impact Term-Based Retrieval Models?
  16. Identify-Conceptualize-Align: A Schema-Adaptive Framework for Unified Entity Recognition and Event Detection
  17. Is a Busy Search Agent a Good One? Overthinking and Overretrieval at Scale
  18. Iterative Structured Pruning for Large Language Models with Multi-Domain Calibration
  19. LongRanker: Efficient One-Pass Document Reranking with Long-Context Large Language Models
  20. Lost in Decomposition: Analyzing and Mitigating the Limitations of Long Context Methods via Context Dependency
  21. Modeling Human-Like Cognition for Stance Detection: Integrating Intuitive Judgment and Analytical Reasoning
  22. One-Pass Decoding for Generative Recommendation with WFST-Constrained A* Search
  23. Projecting Out the Malice: A Global Subspace Approach to LLM Detoxification
  24. RLKD: Distilling LLMs' Reasoning via Reinforcement Learning
  25. Rethinking Implicit Hate Speech Detection: Focusing on Latent Hate Components via Dual-Process Argumentation
  26. RouteRAG: Efficient Retrieval-Augmented Generation from Text and Graph via Reinforcement Learning
  27. Steering Away from Refusal: A Black-box Jailbreak Method Based on First-Token Distribution
  28. Stop Hardening Everything: A Training-Free Neuron-Level Defense for Neural Ranking Models
  29. The Evolution of Thought: Tracking LLM Overthinking via Reasoning Dynamics Analysis
  30. Thinking Forward and Backward: Multi-Objective Reinforcement Learning for Retrieval-Augmented Reasoning
  31. Towards Knowledgeable Deep Research: Framework and Benchmark
  32. Towards Quantitative Summarization Evaluation: An Integrated Atomic-Based Evaluation Framework and Dataset for Text Summarization
  33. a1: Steep Test-time Scaling Law via Environment Augmented Generation
  34. 'I Know You Are Discriminatory!': Automated Substantiating for Individual Fairness Auditing of AI Systems
  35. A Generative Framework for Personalized Sticker Retrieval
  36. A Survey of Link Prediction in N-ary Knowledge Graphs
  37. A Theory for Token-Level Harmonization in Retrieval-Augmented Generation
  38. ALiiCE: Evaluating Positional Fine-grained Citation Generation
  39. Attack-in-the-Chain: Bootstrapping Large Language Models for Attacks Against Black-Box Neural Ranking Models
  40. BLAST: Balanced Sampling Time Series Corpus for Universal Forecasting Models
  41. BSFA: Leveraging the Subspace Dichotomy to Accelerate Neural Network Training
  42. Boosting Retrieval-Augmented Generation with Generation-Augmented Retrieval: A Co-Training Approach
  43. BotTrans: A Multi-source Graph Domain Adaptation Approach for Social Bot Detection
  44. Bridging Queries and Tables through Entities in Open-Domain Table Retrieval
  45. CLIPure: Purification in Latent Space via CLIP for Adversarially Robust Zero-Shot Classification
  46. Can Graph Descriptive Order Affect Solving Graph Problems with LLMs?
  47. Cross-Modal Safety Mechanism Transfer in Large Vision-Language Models
  48. Decoding by Contrasting Knowledge: Enhancing Large Language Model Confidence on Edited Facts
  49. Evaluating Implicit Bias in Large Language Models by Attacking From a Psychometric Perspective
  50. Everything is Editable: Extend Knowledge Editing to Unstructured Data in Large Language Models
  51. Fact-Level Calibration and Correction for Long-Form Generations
  52. Following the Autoregressive Nature of LLM Embeddings via Compression and Alignment
  53. G2S: A General-to-Specific Learning Framework for Temporal Knowledge Graph Forecasting with Large Language Models
  54. Generative Ghost: Investigating Ranking Bias Hidden in AI-Generated Videos
  55. Generative Retrieval for Book Search
  56. Inference-time Alignment in Continuous Space
  57. InfoNCE is a Free Lunch for Semantically guided Graph Contrastive Learning
  58. Is Factuality Enhancement a Free Lunch For LLMs? Better Factuality Can Lead to Worse Context-Faithfulness
  59. Jailbreak LLMs through Internal Stance Manipulation
  60. KnowCoder-X: Boosting Multilingual Information Extraction via Code
  61. Let Topology Speak: Graph Neural Network with Topology-Aware Augmentation
  62. Low-Entropy Watermark Detection via Bayes' Rule Derived Detector
  63. MPRF: Interpretable Stance Detection through Multi-Path Reasoning Framework
  64. MPVStance: Mitigating Hallucinations in Stance Detection with Multi-Perspective Verification
  65. MVAM: Multi-View Attention Method for Fine-Grained Image-Text Matching
  66. NAPPure: Adversarial Purification for Robust Image Classification Under Non-Additive Perturbations
  67. On the Robustness of Generative Information Retrieval Models: An Out-of-Distribution Perspective
  68. On the Scaling of Robustness and Effectiveness in Dense Retrieval
  69. Personalized Denoising Implicit Feedback for Robust Recommender System
  70. QUITO-X: A New Perspective on Context Compression from the Information Bottleneck Theory
  71. Related Knowledge Perturbation Matters: Rethinking Multiple Pieces of Knowledge Editing in Same-Subject
  72. SafetyQuizzer: Timely and Dynamic Evaluation on the Safety of LLMs
  73. Speculative Safety-Aware Decoding
  74. T-MAD: Target-driven Multimodal Alignment for Stance Detection
  75. Tailoring Table Retrieval from a Field-aware Hybrid Matching Perspective
  76. Text2Sql: Pure Fine-Tuning and Pure Knowledge Distillation
  77. The Mirage of Model Editing: Revisiting Evaluation in the Wild
  78. The Silent Saboteur: Imperceptible Adversarial Attacks against Black-Box Retrieval-Augmented Generation Systems
  79. Too Consistent to Detect: A Study of Self-Consistent Errors in LLMs
  80. ToolCoder: A Systematic Code-Empowered Tool Learning Framework for Large Language Models
  81. Towards Event Extraction with Massive Types: LLM-based Collaborative Annotation and Partitioning Extraction
  82. Towards Fully Exploiting LLM Internal States to Enhance Knowledge Boundary Perception
  83. Towards Robust Universal Information Extraction: Dataset, Evaluation, and Solution
  84. Training a Utility-based Retriever Through Shared Context Attribution for Retrieval-Augmented Language Models
  85. Unbiased Learning to Rank with Query-Level Click Propensity Estimation: Beyond Pointwise Observation and Relevance
  86. Utility-Focused LLM Annotation for Retrieval and Retrieval-Augmented Generation
  87. Who is in the Spotlight: The Hidden Bias Undermining Multimodal Retrieval-Augmented Generation
  88. A Multi-Granularity-Aware Aspect Learning Model for Multi-Aspect Dense Retrieval
  89. Adaptive Token Biaser: Knowledge Editing via Biasing Key Entities
  90. Are Large Language Models Good at Utility Judgments?
  91. Blinded by Generated Contexts: How Language Models Merge Generated and Retrieved Contexts When Knowledge Conflicts?
  92. Bootstrapped Pre-training with Dynamic Identifier Prediction for Generative Retrieval
  93. CausalDiff: Causality-Inspired Disentanglement via Diffusion Model for Adversarial Defense
  94. Classifier Guidance Enhances Diffusion-Based Adversarial Purification by Preserving Predictive Information
  95. Controlling Risk of Retrieval-augmented Generation: A Counterfactual Prompting Framework
  96. Enhancing Training Data Attribution for Large Language Models with Fitting Error Consideration
  97. FCS-HGNN: Flexible Multi-type Community Search in Heterogeneous Information Networks
  98. GaQR: An Efficient Generation-augmented Question Rewriter
  99. Generative Retrieval Meets Multi-Graded Relevance
  100. Graph Summarization for Preserving Spectral Characteristics
  101. Improving the Shortest Plank: Vulnerability-Aware Adversarial Training for Robust Recommender System
  102. Invisible Relevance Bias: Text-Image Retrieval Models Prefer AI-Generated Images
  103. KnowCoder: Coding Structured Knowledge into LLMs for Universal Information Extraction
  104. LINKAGE: Listwise Ranking among Varied-Quality References for Non-Factoid QA Evaluation via LLMs
  105. LPNL: Scalable Link Prediction with Large Language Models
  106. List-aware Reranking-Truncation Joint Model for Search and Retrieval-augmented Generation
  107. LoRec: Combating Poisons with Large Language Model for Robust Sequential Recommendation
  108. Look Globally and Reason: Two-stage Path Reasoning over Sparse Knowledge Graphs
  109. MORE: Multi-mOdal REtrieval Augmented Generative Commonsense Reasoning
  110. Matching Knowledge Graphs in Entity Embedding Spaces: An Experimental Study [Extended Abstract]
  111. Multi-granular Adversarial Attacks against Black-box Neural Ranking Models
  112. Negative as Positive: Enhancing Out-of-distribution Generalization for Graph Contrastive Learning
  113. PDE+: Enhancing Generalization via PDE with Adaptive Distributional Diffusion
  114. Perturbation-Invariant Adversarial Training for Neural Ranking Models: Improving the Effectiveness-Robustness Trade-Off
  115. Plot Retrieval as an Assessment of Abstract Semantic Association
  116. Pretraining Data Detection for Large Language Models: A Divergence-based Calibration Method
  117. RoCEL: Advancing Table Entity Linking through Distinctive Row and Column Contexts
  118. SACH: Significant-Attributed Community Search in Heterogeneous Information Networks
  119. SLANG: New Concept Comprehension of Large Language Models
  120. Search-in-the-Chain: Interactively Enhancing Large Language Models with Search for Knowledge-intensive Tasks
  121. TEA: Test-Time Energy Adaptation
  122. The Butterfly Effect of Model Editing: Few Edits Can Trigger Large Language Models Collapse
  123. Think Before You Speak: Cultivating Communication Skills of Large Language Models via Inner Monologue
  124. Understanding and Improving Adversarial Collaborative Filtering for Robust Recommendation
  125. Unsupervised Information Refinement Training of Large Language Models for Retrieval-Augmented Generation
  126. When Do LLMs Need Retrieval Augmentation? Mitigating LLMs' Overconfidence Helps Retrieval Augmentation