PPaperPicks

Bing Qin

Harbin Institute of Technology, Research Center for Social Computing and Information Retrieval, China

119 papers at tracked venues · 94 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Adaptive Backtracking for Privacy Protection in Large Language Models
  2. Beyond the Flat Sequence: Hierarchical and Preference-Aware Generative Recommendations
  3. Breaking Language Preference in Multilingual RAG via Language-Controllable Retrieval and Language-Agnostic Reasoning
  4. CARE-Bench: A Benchmark of Diverse Client Simulations Guided by Expert Principles for Evaluating LLMs in Psychological Counseling
  5. Collaborative Chain-of-Agents for Parametric-Retrieved Knowledge Synergy
  6. Consolidation or Adaptation? PRISM: Disentangling SFT and RL Data via Gradient Concentration
  7. Culture-Aware Machine Translation in Large Language Models: Benchmarking and Investigation
  8. CultureRL: Internalizing Cultural Principles in Large Language Models via Norm-Driven Reinforcement Learning
  9. EMTIR-GRPO: Efficient Multi-Tool Augmented Large Language Models via Reinforcement Learning
  10. ENPMR-Bench: Benchmarking Proactive Memory Retrieval for Emotional Support Agents
  11. Fine-Mem: Fine-Grained Feedback Alignment for Long-Horizon Memory Management
  12. From Implicit Graph Encoding to Explicit Evidence: A Training-Free LLM Framework for Temporal Knowledge Graph Reasoning
  13. Graph Reasoning Paradigm: Structured and Symbolic Reasoning with Topology-Aware Reinforcement Learning for Large Language Models
  14. LangGPS: Language Separability Guided Data Pre-Selection for Joint Multilingual Instruction Tuning
  15. Large Language Models Are Still Misled by Simple Bias Ensembles
  16. MAESTRO: Meta-learning Adaptive Estimation of Scalarization Trade-offs for Reward Optimization
  17. MCGA: A Multi-task Classical Chinese Literary Genre Audio Corpus
  18. MDC-Bench: A Multidisciplinary Causal Benchmark Based on Causal Structures for Evaluating Large Language Models
  19. MPR-GUI: Benchmarking and Enhancing Multilingual Perception and Reasoning in GUI Agents
  20. On Safety Risks in Experience-Driven Self-Evolving Agents
  21. PARIF: Pushing the Pareto Frontier of Instruction Following and Reasoning with Curriculum Reinforcement Learning
  22. Psychological Counseling Cannot Be Achieved Overnight: Automated Psychological Counseling Through Multi-Session Conversations
  23. Question Tells You Where the Answer Is: Intention-aware Long-Context KV Cache Compression
  24. SAVOIR: Learning Social Savoir-Faire via Shapley-based Reward Attribution
  25. Stratagem: Learning Transferable Reasoning via Trajectory-Modulated Game Self-Play
  26. TEA-Bench: A Systematic Benchmarking of Tool-enhanced Emotional Support Dialogue Agent
  27. Toward Secure Tuning: Mitigating Security Risks from Instruction Fine-Tuning
  28. Trade-offs in Large Reasoning Models: An Empirical Analysis of Deliberative and Adaptive Reasoning over Foundational Capabilities
  29. Two-Stage Parameter Alignment for Multi-LoRA Merging in Large Language Models
  30. Unlocking Multilingual Reasoning Capability of LLMs and LVLMs through Representation Engineering
  31. WebAnchor: Anchoring Agent Planning to Stabilize Long-Horizon Web Reasoning
  32. When Correct Beliefs Collapse: Epistemic Resilience of LLMs under Clinical Pressure
  33. When Personalization Legitimizes Risks: Uncovering Safety Vulnerabilities in Personalized Dialogue Agents
  34. x1: Learning to Think Adaptively Across Languages and Cultures
  35. AdaSteer: Your Aligned LLM is Inherently an Adaptive Jailbreak Defender
  36. Alleviating Hallucinations from Knowledge Misalignment in Large Language Models via Selective Abstention Learning
  37. AnRe: Analogical Replay for Temporal Knowledge Graph Forecasting
  38. Analyzing the Rapid Generalization of SFT via the Perspective of Attention Head Activation Patterns
  39. Balancing Forget Quality and Model Utility: A Reverse KL-Divergence Knowledge Distillation Approach for Better Unlearning in LLMs
  40. Beware of Your Po! Measuring and Mitigating AI Safety Risks in Role-Play Fine-Tuning of LLMs
  41. Beyond Fixed Length: Bucket Pre-training is All You Need
  42. Beyond Frameworks: Unpacking Collaboration Strategies in Multi-Agent Systems
  43. Beyond Similarity: A Gradient-based Graph Method for Instruction Tuning Data Selection
  44. Beyond Snapshots: A Multimodal User-Level Dataset for Depression Detection in Dynamic Social Media Streams
  45. Breaking the Reasoning Barrier A Survey on LLM Complex Reasoning through the Lens of Self-Evolution
  46. CC-Tuning: A Cross-Lingual Connection Mechanism for Improving Joint Multilingual Supervised Fine-Tuning
  47. CLAIM: Mitigating Multilingual Object Hallucination in Large Vision-Language Models with Cross-Lingual Attention Intervention
  48. Chain of Strategy Optimization Makes Large Language Models Better Emotional Supporter
  49. Com² : A Causal-Guided Benchmark for Exploring Complex Commonsense Reasoning in Large Language Models
  50. Context-Aware Hierarchical Taxonomy Generation for Scientific Papers via LLM-Guided Multi-Aspect Clustering
  51. Cross-Lingual Text-Rich Visual Comprehension: An Information Theory Perspective
  52. EffiVLM-BENCH: A Comprehensive Benchmark for Evaluating Training-Free Acceleration in Large Vision-Language Models
  53. End-to-End Learnable Psychiatric Scale Guided Risky Post Screening for Depression Detection on Social Media
  54. Enhancing Non-English Capabilities of English-Centric Large Language Models Through Deep Supervision Fine-Tuning
  55. Ensuring Responses Contain Appropriate Images: Timing Judgment for Multimodal Responses
  56. ExpeTrans: LLMs Are Experiential Transfer Learners
  57. Exploring Large Language Models for Effective Rumor Detection on Social Media
  58. FroM: Frobenius Norm-Based Data-Free Adaptive Model Merging
  59. From Hypothesis to Publication: A Comprehensive Survey of AI-Driven Research Support Systems
  60. From Specific-MLLMs to Omni-MLLMs: A Survey on MLLMs Aligned with Multi-modalities
  61. GainRAG: Preference Alignment in Retrieval-Augmented Generation through Gain Signal Synthesis
  62. How Does Sequence Modeling Architecture Influence Base Capabilities of Pre-trained Language Models? Exploring Key Architecture Design Principles to Avoid Base Capabilities Degradation
  63. How do Language Models Reshape Entity Alignment? A Survey of LM-Driven EA Methods: Advances, Benchmarks, and Future
  64. Improved Diffusion-based Generative Model with Better Adversarial Robustness
  65. Improving Contextual Faithfulness of Large Language Models via Retrieval Heads-Induced Optimization
  66. Investigating and Enhancing the Robustness of Large Multimodal Models Against Temporal Inconsistency
  67. Investigating the Security Threat Arising from "Yes-No" Implicit Bias in Large Language Models
  68. Length Controlled Generation for Black-box LLMs
  69. Look Beyond Feeling: Unveiling Latent Needs from Implicit Expressions for Proactive Emotional Support
  70. MA-GTS: A Multi-Agent Framework for Solving Complex Graph Problems in Real-World Applications
  71. MPO: Multilingual Safety Alignment via Reward Gap Optimization
  72. Making LLMs Better Many-to-Many Speech-to-Text Translators with Curriculum Learning
  73. Natural Logic at the Core: Dynamic Rewards for Entailment Tree Generation
  74. One for All: Update Parameterized Knowledge Across Multiple Models with Once Edit
  75. Ontology-Guided Reverse Thinking Makes Large Language Models Stronger on Knowledge Graph Question Answering
  76. Probing and Boosting Large Language Models Capabilities via Attention Heads
  77. Self-Critique Guided Iterative Reasoning for Multi-hop Question Answering
  78. Separate the Wheat from the Chaff: A Post-Hoc Approach to Safety Re-Alignment for Fine-Tuned Language Models
  79. Simulating Before Planning: Constructing Intrinsic User World Model for User-Tailored Dialogue Policy Planning
  80. Simulation-Free Hierarchical Latent Policy Planning for Proactive Dialogues
  81. Teaching Language Models to Evolve with Users: Dynamic Profile Modeling for Personalized Alignment
  82. Tool Zero: Training Tool-Augmented LLMs via Pure RL from Scratch
  83. UFO-RL: Uncertainty-Focused Optimization for Efficient Reinforcement Learning Data Selection
  84. When Less Language is More: Language-Reasoning Disentanglement Makes LLMs Better Multilingual Reasoners
  85. iTool: Reinforced Fine-Tuning with Dynamic Deficiency Calibration for Advanced Tool Use
  86. AS-ES Learning: Towards efficient CoT learning in small models
  87. Advancing Large Language Model Attribution through Self-Improving
  88. Aligning Translation-Specific Understanding to General Understanding in Large Language Models
  89. An Information Bottleneck Perspective for Effective Noise Filtering on Retrieval-Augmented Generation
  90. BC-Prover: Backward Chaining Prover for Formal Theorem Proving
  91. BeamAggR: Beam Aggregation Reasoning over Multi-source Knowledge for Multi-hop Question Answering
  92. Both Matter: Enhancing the Emotional Intelligence of Large Language Models without Compromising the General Intelligence
  93. Causal-Guided Active Learning for Debiasing Large Language Models
  94. CogGPT: Unleashing the Power of Cognitive Dynamics on Large Language Models
  95. Deciphering the Impact of Pretraining Data on Large Language Models through Machine Unlearning
  96. Decomposing Argumentative Essay Generation via Dialectical Planning of Complex Reasoning
  97. Discrete Modeling via Boundary Conditional Diffusion Processes
  98. Divide-and-Conquer Meets Consensus: Unleashing the Power of Functions in Code Generation
  99. Ensemble Learning for Heterogeneous Large Language Models with Deep Parallel Collaboration
  100. Extending Context Window of Large Language Models from a Distributional Perspective
  101. From Artificially Real to Real: Leveraging Pseudo Data from Large Language Models for Low-Resource Molecule Discovery
  102. GUIDE: A Guideline-Guided Dataset for Instructional Video Comprehension
  103. GlobeSumm: A Challenging Benchmark Towards Unifying Multi-lingual, Cross-lingual and Multi-document News Summarization
  104. How does Architecture Influence the Base Capabilities of Pre-trained Language Models? A Case Study Based on FFN-Wider and MoE Transformers
  105. Improving In-Context Learning with Prediction Feedback for Sentiment Analysis
  106. Infrared-LLaVA: Enhancing Understanding of Infrared Images in Multi-Modal Large Language Models
  107. Investigating and Mitigating the Multimodal Hallucination Snowballing in Large Vision-Language Models
  108. Learning Fine-Grained Grounded Citations for Attributed Large Language Models
  109. Length Extrapolation of Transformers: A Survey from the Perspective of Positional Encoding
  110. Manifold-Based Verbalizer Space Re-embedding for Tuning-Free Prompt-Based Classification
  111. Meaningful Learning: Enhancing Abstract Reasoning in Large Language Models via Generic Fact Guidance
  112. MoGU: A Framework for Enhancing Safety of LLMs While Preserving Their Usability
  113. MolTailor: Tailoring Chemical Molecular Representation to Specific Tasks via Text Prompts
  114. Navigate through Enigmatic Labyrinth A Survey of Chain of Thought Reasoning: Advances, Frontiers and Future
  115. Planning Like Human: A Dual-process Framework for Dialogue Planning
  116. SAPT: A Shared Attention Framework for Parameter-Efficient Continual Learning of Large Language Models
  117. Self-Evolving GPT: A Lifelong Autonomous Experiential Learner
  118. TimeBench: A Comprehensive Evaluation of Temporal Reasoning Abilities in Large Language Models
  119. Towards Benchmarking Situational Awareness of Large Language Models: Comprehensive Benchmark, Evaluation and Analysis