PPaperPicks

Qingyao Ai

Tsinghua University, Beijing, China

58 papers at tracked venues · 47 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Analytical Search
  2. Auto-PRE: An Automatic and Cost-Efficient Peer-Review Framework for Language Generation Evaluation
  3. Beyond Experience Retrieval: Learning to Generate Utility-Optimized Structured Experience for Frozen LLMs
  4. Chinese Court Simulation with LLM-Based Agents System
  5. Enhancing Judgment Document Generation via Agentic Legal Information Collection and Rubric-Guided Optimization
  6. Equity vs. Equality: Optimizing Ranking Fairness for Tailored Provider Needs
  7. Generalized Pseudo-Relevance Feedback
  8. Investigating Prosocial Behavior Theory in LLM Agents Under Policy-Induced Inequities
  9. Joint Evaluation of Answer and Reasoning Consistency for Hallucination Detection in Large Reasoning Models
  10. LegalKit: A Modular Toolkit for Efficient Legal AI Evaluation
  11. Simulating Dispute Mediation with LLM-Based Agents for Legal Research
  12. SurGE: A Benchmark and Evaluation Framework for Scientific Survey Generation
  13. TEC: A Collection of Human Trial-and-error Trajectories for Problem Solving
  14. TwinVoice: A Multi-dimensional Benchmark Towards Digital Twins via LLM Persona Simulation
  15. Unsupervised Dense Retrieval with Conterfactual Contrastive Learning
  16. Augmenting Multi-Agent Communication with State Delta Trajectory
  17. BLADE: Enhancing Black-Box Large Language Models with Small Domain-Specific Models
  18. Brain Image Reconstruction with Retrieval-Augmented Diffusion
  19. CalibraEval: Calibrating Prediction Distribution to Mitigate Selection Bias in LLMs-as-Judges
  20. DELTA: Pre-Train a Discriminative Encoder for Legal Case Retrieval via Structural Word Alignment
  21. Decoupling Knowledge and Context: An Efficient and Effective Retrieval Augmented Generation Framework via Cross Attention
  22. Decoupling Reasoning and Knowledge Injection for In-Context Knowledge Editing
  23. Dynamic and Parametric Retrieval-Augmented Generation
  24. Investigating the Robustness of Counterfactual Learning to Rank Models: A Reproducibility Study
  25. JuDGE: Benchmarking Judgment Document Generation for Chinese Legal System
  26. JustEva: A Toolkit to Evaluate LLM Fairness in Legal Knowledge Inference
  27. Knowledge Editing through Chain-of-Thought
  28. Learning LLM-as-a-Judge for Preference Alignment
  29. LegalAgentBench: Evaluating LLM Agents in Legal Domain
  30. LexRAG: Benchmarking Retrieval-Augmented Generation in Multi-Turn Legal Consultation Conversation
  31. PEPE: Long-context Extension for Large Language Models via Periodic Extrapolation Positional Encodings
  32. Parametric Retrieval Augmented Generation
  33. Qilin: A Multimodal Information Retrieval Dataset with APP-level User Sessions
  34. Robust Fine-tuning for Retrieval Augmented Generation against Retrieval Defects
  35. SelfRACG: Enabling LLMs to Self-Express and Retrieve for Code Generation
  36. SimVBG: Simulating Individual Values by Backstory Generation
  37. The 1st NIP@IR Workshop on New Interaction Paradigms for Information Retrieval in the Era of Generative AI
  38. Understanding the Effect of Opinion Polarization in Short Video Browsing
  39. An In-depth Investigation of User Response Simulation for Conversational Search
  40. Automatic Large Language Model Evaluation via Peer Review
  41. Capability-aware Prompt Reformulation Learning for Text-to-Image Generation
  42. Combining Multiple Supervision for Robust Zero-Shot Dense Retrieval
  43. DRAGIN: Dynamic Retrieval Augmented Generation based on the Real-time Information Needs of Large Language Models
  44. EEG-SVRec: An EEG Dataset with User Multidimensional Affective Engagement Labels in Short Video Recommendation
  45. GNN4EEG: A Benchmark and Toolkit for Electroencephalography Classification with Graph Neural Network
  46. LeCaRDv2: A Large-Scale Chinese Legal Case Retrieval Dataset
  47. LeDQA: A Chinese Legal Case Document-based Question Answering Dataset
  48. LexEval: A Comprehensive Chinese Legal Benchmark for Evaluating Large Language Models
  49. Mitigating Exploitation Bias in Learning to Rank with an Uncertainty-aware Empirical Bayes Approach
  50. Prompt Refinement with Image Pivot for Text-to-Image Generation
  51. Query Augmentation with Brain Signals
  52. STARD: A Chinese Statute Retrieval Dataset Derived from Real-life Queries by Non-professionals
  53. Scaling Laws For Dense Retrieval
  54. Sequential Recommendation with Latent Relations based on Large Language Model
  55. Unbiased Learning-to-Rank Needs Unconfounded Propensity Estimation
  56. Unsupervised Large Language Model Alignment for Information Retrieval via Contrastive Feedback
  57. Unsupervised Real-Time Hallucination Detection based on the Internal States of Large Language Models
  58. Wikiformer: Pre-training with Structured Information of Wikipedia for Ad-Hoc Retrieval