PPaperPicks

Jeff Z. Pan

The University of Edinburgh, UK

49 papers at tracked venues · 33 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Beyond Timestamps: Bridging Forward and Backward Reasoning in Temporal Numerical and Relational Understanding
  2. Illusions of Confidence? Diagnosing LLM Truthfulness via Neighborhood Consistency
  3. MINTQA: A Multi-Hop Question Answering Benchmark for Evaluating LLMs on New and Long-tail Knowledge
  4. MemGuide: Intent-Driven Memory Selection for Goal-Oriented Multi-Session LLM Agents
  5. Memory-R1: Enhancing Large Language Model Agents to Manage and Utilize Memories via Reinforcement Learning
  6. Mitigating Context Interference for Reliable and Efficient Search Agents
  7. ODL-TempLLM: Ontology-Guided and Description Logic-Reasoned Temporal Reasoning with LLMs
  8. Plan Then Retrieve: Reinforcement Learning-Guided Complex Reasoning over Knowledge Graphs
  9. ReLUPruner: Rethinking ReLU Importance with Taylor Expansion for Efficient Private Inference
  10. Self-Sum: Teaching an Agent to Decide Itself When and What to Summarize
  11. Semantic Alignment of Malicious Question Based on Contrastive Semantic Networks and Data Augmentation (Abstract Reprint)
  12. TaxReasoning: Benchmarking Knowledge-Intensive Mathematical Reasoning with Evolving Tax Laws
  13. Uncovering and Mitigating Transient Blindness in Multimodal Model Editing
  14. A Controllable Examination for Long-Context Language Models
  15. Can LLMs Evaluate Complex Attribution in QA? Automatic Benchmarking using Knowledge Graphs
  16. Dark Side of Modalities: Reinforced Multimodal Distillation for Multimodal Knowledge Graph Reasoning
  17. Decomposing and Revising What Language Models Generate
  18. Don't Forget the Base Retriever! A Low-Resource Graph-based Retriever for Multi-hop Question Answering
  19. Evaluating and Improving Graph to Text Generation with Large Language Models
  20. From an LLM Swarm to a PDDL-empowered Hive: Planning Self-executed Instructions in a Multi-modal Jungle
  21. GeAR: Graph-enhanced Agent for Retrieval-augmented Generation
  22. GenTool: Enhancing Tool Generalization in Language Models through Zero-to-One and Weak-to-Strong Simulation
  23. LLM Shots: Best Fired at System or User Prompts?
  24. Long-Form Information Alignment Evaluation Beyond Atomic Facts
  25. Masking in Multi-hop QA: An Analysis of How Language Models Perform with Context Permutation
  26. MiCEval: Unveiling Multimodal Chain of Thought's Quality via Image Description and Reasoning Steps
  27. Multi-level Matching Network for Multimodal Entity Linking
  28. Multi-level Mixture of Experts for Multimodal Entity Linking
  29. ReSearch: Learning to Reason with Search for LLMs via Reinforcement Learning
  30. Rethinking Stateful Tool Use in Multi-Turn Dialogues: Benchmarks and Challenges
  31. Schema-Constrained Grammar-Guided Generation of GQL Queries from Natural Language
  32. Self-Reasoning Language Models: Unfold Hidden Reasoning Chains with Few Reasoning Catalyst
  33. TS-SQL: Test-driven Self-refinement for Text-to-SQL
  34. [PromptEng] Second International Workshop on Prompt Engineering for Pre-Trained Language Models
  35. A Large-scale Offer Alignment Model for Partitioning Filtering and Matching Product Offers
  36. A Usage-centric Take on Intent Understanding in E-Commerce
  37. An Empirical Study on Parameter-Efficient Fine-Tuning for MultiModal Large Language Models
  38. AppBench: Planning of Multiple APIs from Various APPs for Complex User Instruction
  39. CoTKR: Chain-of-Thought Enhanced Knowledge Rewriting for Complex Knowledge Graph Question Answering
  40. Empowering Large Language Models: Tool Learning for Real-World Interaction
  41. Improving Retrieval-augmented Text-to-SQL with AST-based Ranking and Schema Pruning
  42. Inference Helps PLMs' Conceptual Understanding: Improving the Abstract Inference Ability with Hierarchical Conceptual Entailment Graphs
  43. InstructEd: Soft-Instruction Tuning for Model Editing with Hops
  44. InstructIE: A Bilingual Instruction-based Information Extraction Dataset
  45. Knowledge-Aware Neuron Interpretation for Scene Classification
  46. Learning to Plan for Retrieval-Augmented Large Language Models from Knowledge Graphs
  47. Less is More: Making Smaller Language Models Competent Subgraph Retrievers for Multi-hop KGQA
  48. UniArk: Improving Generalisation and Consistency for Factual Knowledge Extraction through Debiasing
  49. [PromptEng] First International Workshop on Prompt Engineering for Pre-Trained Language Models