PPaperPicks

Yangqiu Song

Hong Kong University of Science and Technology (HKUST), Department of Computer Science and Engineering, Tai Po Tsai, Hong Kong

71 papers at tracked venues · 45 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AutoGraph-R1: End-to-End Reinforcement Learning for Knowledge Graph Construction
  2. AutoSchemaKG: Autonomous Knowledge Graph Construction through Dynamic Schema Induction from Web-Scale Corpora
  3. ContextLens: Modeling Imperfect Privacy and Safety Context for Legal Compliance
  4. DIXITWORLD: Evaluating Multimodal Abductive Reasoning in Vision-Language Models with Multi-Agent Dixit Gameplay
  5. DeepPlanner: Scaling Planning Capability for Deep Research Agents via Advantage Shaping
  6. GrandGuard: Taxonomy, Benchmark, and Safeguards for Elderly-Chatbot Interaction Safety
  7. InferenceDynamics: Adaptive LLM Routing through Structured Capability and Knowledge Profiling
  8. Intention Knowledge Graph Construction for User Intention Relation Modeling
  9. Into the Gray Zone: Domain Contexts Can Blur LLM Safety Boundaries
  10. OmniCompliance-100K: A Multi-Domain, Rule-Grounded, Real-World Safety Compliance Dataset
  11. Robustness via Referencing: Defending against Prompt Injection Attacks by Referencing the Executed Instruction
  12. SessionIntentBench: A Multi-task Inter-session Intention-shift Modeling Benchmark for E-commerce Customer Behavior Understanding
  13. Unifying Deductive and Abductive Reasoning in Knowledge Graphs with Masked Diffusion Model
  14. XToM: Exploring the Multilingual Theory of Mind for Large Language Models
  15. arXiv2Table: Toward Realistic Benchmarking and Evaluation for LLM-Based Literature-Review Table Generation
  16. A Survey of RAG-Reasoning Systems in Large Language Models
  17. Backdoor-Powered Prompt Injection Attacks Nullify Defense Methods
  18. Can Indirect Prompt Injection Attacks Be Detected and Removed?
  19. Chain of Attack: On the Robustness of Vision-Language Models Against Transfer-Based Adversarial Attacks
  20. ComparisonQA: Evaluating Factuality Robustness of LLMs Through Knowledge Frequency Control and Uncertainty
  21. ConKE: Conceptualization-Augmented Knowledge Editing in Large Language Models for Commonsense Reasoning
  22. Concept-Reversed Winograd Schema Challenge: Evaluating and Improving Robust Reasoning in Large Language Models via Abstraction
  23. Context Reasoner: Incentivizing Reasoning Capability for Contextualized Privacy and Safety Compliance via Reinforcement Learning
  24. Defense Against Prompt Injection Attack by Leveraging Attack Techniques
  25. DivScene: Towards Open-Vocabulary Object Navigation with Large Vision Language Models in Diverse Scenes
  26. EFOk-CQA: Towards Knowledge Graph Complex Query Answering beyond Set Operation
  27. EcomScriptBench: A Multi-task Benchmark for E-commerce Script Planning via Step-wise Intention-Driven Product Association
  28. Enhancing Transformers for Generalizable First-Order Logical Entailment
  29. Extending Complex Logical Queries on Uncertain Knowledge Graphs
  30. From Automation to Autonomy: A Survey on Large Language Models in Scientific Discovery
  31. Frontiers in Graph Machine Learning for the Large Model Era
  32. InteGround: On the Evaluation of Verification and Retrieval Planning in Integrative Grounding
  33. KnowShiftQA: How Robust are RAG Systems when Textbook Knowledge Shifts in K-12 Education?
  34. LogiDynamics: Unraveling the Dynamics of Inductive, Abductive and Deductive Logical Inferences in LLM Reasoning
  35. MARS: Benchmarking the Metaphysical Reasoning Abilities of Language Models with a Multi-task Evaluation Dataset
  36. MCIP: Protecting MCP Safety via Model Contextual Integrity Protocol
  37. MMLongBench: Benchmarking Long-Context Vision-Language Models Effectively and Thoroughly
  38. On the Role of Entity and Event Level Conceptualization in Generalizable Reasoning: A Survey of Tasks, Methods, Applications, and Future Directions
  39. Patterns Over Principles: The Fragility of Inductive Reasoning in LLMs under Noisy Observations
  40. PrivaCI-Bench: Evaluating Privacy with Contextual Integrity and Legal Compliance
  41. Privacy Checklist: Privacy Violation Detection Grounding on Contextual Integrity Theory
  42. Revisiting Epistemic Markers in Confidence Estimation: Can Markers Accurately Reflect Large Language Models' Uncertainty?
  43. Revolve: Optimizing AI Systems by Tracking Response Evolution in Textual Optimization
  44. Simulate and Eliminate: Revoke Backdoors for Generative Large Language Models
  45. SwitchLingua: The First Large-Scale Multilingual and Multi-Ethnic Code-Switching Dataset
  46. TopicAttack: An Indirect Prompt Injection Attack via Topic Transition
  47. AbsInstruct: Eliciting Abstraction Ability from LLMs through Explanation Tuning with Plausibility Estimation
  48. AbsPyramid: Benchmarking the Abstraction Ability of Language Models with a Unified Entailment Graph
  49. ActPlan-1K: Benchmarking the Procedural Planning Ability of Visual Language Models in Household Activities
  50. Advancing Abductive Reasoning in Knowledge Graphs through Complex Logical Hypothesis Generation
  51. Audience Persona Knowledge-Aligned Prompt Tuning Method for Online Debate
  52. CANDLE: Iterative Conceptualization and Instantiation Distillation from Large Language Models for Commonsense Reasoning
  53. COSMO: A Large-Scale E-commerce Common Sense Knowledge Generation and Serving System at Amazon
  54. Complex Reasoning over Logical Queries on Commonsense Knowledge Graphs
  55. ECON: On the Detection and Resolution of Evidence Conflicts
  56. Generate-on-Graph: Treat LLM as both Agent and KG for Incomplete Knowledge Graph Question Answering
  57. Getting Sick After Seeing a Doctor? Diagnosing and Mitigating Knowledge Conflicts in Event Temporal Reasoning
  58. GoldCoin: Grounding Large Language Models in Privacy Laws via Contextual Integrity Theory
  59. IntentionQA: A Benchmark for Evaluating Purchase Intention Comprehension Abilities of Language Models in E-commerce
  60. MIND: Multimodal Shopping Intention Distillation from Large Vision-language Models for E-commerce Purchase Understanding
  61. Miko: Multimodal Intention Knowledge Distillation from Large Language Models for Social-Media Commonsense Discovery
  62. NegotiationToM: A Benchmark for Stress-testing Machine Theory of Mind on Negotiation Surrounding
  63. New Frontiers of Knowledge Graph Reasoning: Recent Advances and Future Trends
  64. PrivLM-Bench: A Multi-level Privacy Evaluation Benchmark for Language Models
  65. Privacy-Preserved Neural Graph Databases
  66. Rethinking Complex Queries on Knowledge Graphs with Neural Link Predictors
  67. Rethinking the Bounds of LLM Reasoning: Are Multi-Agent Discussions the Key?
  68. Text-Tuple-Table: Towards Information Integration in Text-to-Table Generation via Global Tuple Extraction
  69. Understanding Inter-Session Intentions via Complex Logical Reasoning
  70. UrbanCross: Enhancing Satellite Image-Text Retrieval with Cross-Domain Adaptation
  71. User Consented Federated Recommender System Against Personalized Attribute Inference Attack