P
PaperPicks
Conferences
Benjamin Van Durme
Johns Hopkins University, USA
54 papers at tracked venues · 31 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID 0000-0003-4328-4288 ↗
Google Scholar ↗
Homepage ↗
Venues
ACL
×12
EMNLP
×10
NAACL
×9
ICLR
×6
AAAI
×5
CVPR
×3
SIGIR
×3
ECIR
×2
IJCAI
×2
ICML
×1
NeurIPS
×1
Frequent coauthors
Kate Sanders
DBLP profile ↗
ORCID search ↗
×4
Jingyu Zhang
DBLP profile ↗
ORCID search ↗
×4
Nathaniel Weir
DBLP profile ↗
ORCID search ↗
×4
Zhengping Jiang
DBLP profile ↗
ORCID search ↗
×3
Orion Weller
DBLP profile ↗
ORCID search ↗
×3
Saron Samuel
DBLP profile ↗
ORCID search ↗
×2
William Jurayj
DBLP profile ↗
ORCID search ↗
×2
Jiefu Ou
DBLP profile ↗
ORCID search ↗
×2
Abe Bohan Hou
DBLP profile ↗
ORCID search ↗
×2
Harsh Jhamtani
DBLP profile ↗
ORCID search ↗
×2
Dongwei Jiang
DBLP profile ↗
ORCID search ↗
×2
Andrew Blair-Stanek
DBLP profile ↗
ORCID search ↗
×1
Papers
Bonsai: Interpretable Tree-Adaptive Grounded Reasoning
AAAI 2026
·
Kate Sanders
DBLP profile ↗
ORCID search ↗
Can LLMs Identify Tax Abuse?
AAAI 2026
·
Andrew Blair-Stanek
DBLP profile ↗
ORCID search ↗
CoverageBench: Evaluating Information Coverage across Tasks and Domains
SIGIR 2026
·
Saron Samuel
DBLP profile ↗
ORCID search ↗
Does Reasoning Make Search More Fair? Comparing Fairness in Reasoning and Non-reasoning Rerankers
ECIR 2026
·
Saron Samuel
DBLP profile ↗
ORCID search ↗
How Grounded is Wikipedia? A Study on Structured Evidential Support and Retrieval
ACL 2026
·
William Gantt Walden
DBLP profile ↗
ORCID search ↗
Language Models and Logic Programs for Trustworthy Tax Reasoning
AAAI 2026
·
William Jurayj
DBLP profile ↗
ORCID search ↗
Multi-Vector Index Compression in Any Modality
SIGIR 2026
·
Hanxiang Qin
DBLP profile ↗
ORCID search ↗
WikiVideo: Article Generation from Multiple Videos
ACL 2026
·
Alexander Martin
DBLP profile ↗
ORCID search ↗
arXiv2Table: Toward Realistic Benchmarking and Evaluation for LLM-Based Literature-Review Table Generation
ACL 2026
·
Weiqi Wang
DBLP profile ↗
ORCID search ↗
CLAIMCHECK: How Grounded are LLM Critiques of Scientific Papers?
EMNLP 2025
·
Jiefu Ou
DBLP profile ↗
ORCID search ↗
CLERC: A Dataset for U. S. Legal Case Retrieval and Retrieval-Augmented Analysis Generation
NAACL 2025
·
Abe Bohan Hou
DBLP profile ↗
ORCID search ↗
Certified Mitigation of Worst-Case LLM Copyright Infringement
EMNLP 2025
·
Jingyu Zhang
DBLP profile ↗
ORCID search ↗
Conformal Linguistic Calibration: Trading-off between Factuality and Specificity
NeurIPS 2025
·
Zhengping Jiang
DBLP profile ↗
ORCID search ↗
Controllable Safety Alignment: Inference-Time Adaptation to Diverse Safety Requirements
ICLR 2025
·
Jingyu Zhang
DBLP profile ↗
ORCID search ↗
Core: Robust Factual Precision with Informative Sub-Claim Identification
ACL 2025
·
Zhengping Jiang
DBLP profile ↗
ORCID search ↗
DnDScore: Decontextualization and Decomposition for Factuality Verification in Long-Form Text Generation
EMNLP 2025
·
Miriam Wanner
DBLP profile ↗
ORCID search ↗
FollowIR: Evaluating and Teaching Information Retrieval Models to Follow Instructions
NAACL 2025
·
Orion Weller
DBLP profile ↗
ORCID search ↗
From Models to Microtheories: Distilling a Model's Topical Knowledge for Grounded Question-Answering
ICLR 2025
·
Nathaniel Weir
DBLP profile ↗
ORCID search ↗
Generative Adapter: Contextualizing Language Models in Parameters with A Single Forward Pass
ICLR 2025
·
Tong Chen
DBLP profile ↗
ORCID search ↗
Is That Your Final Answer? Test-Time Scaling Improves Selective Question Answering
ACL 2025
·
William Jurayj
DBLP profile ↗
ORCID search ↗
Jailbreak Distillation: Renewable Safety Benchmarking
EMNLP 2025
·
Jingyu Zhang
DBLP profile ↗
ORCID search ↗
LLM Agents for Coordinating Multi-User Information Gathering
ACL 2025
·
Harsh Jhamtani
DBLP profile ↗
ORCID search ↗
MICE for CATs: Model-Internal Confidence Estimation for Calibrating Agents with Tools
NAACL 2025
·
Nishant Subramani
DBLP profile ↗
ORCID search ↗
Multi-Field Adaptive Retrieval
ICLR 2025
·
Millicent Li
DBLP profile ↗
ORCID search ↗
MultiVENT 2.0: A Massive Multilingual Benchmark for Event-Centric Video Retrieval
CVPR 2025
·
Reno Kriz
DBLP profile ↗
ORCID search ↗
Promptriever: Instruction-Trained Retrievers Can Be Prompted Like Language Models
ICLR 2025
·
Orion Weller
DBLP profile ↗
ORCID search ↗
RATIONALYST: Pre-training Process-Supervision for Improving Reasoning
ACL 2025
·
Dongwei Jiang
DBLP profile ↗
ORCID search ↗
RE-AdaptIR: Improving Information Retrieval through Reverse Engineered Adaptation
SIGIR 2025
·
William Fleshman
DBLP profile ↗
ORCID search ↗
SELF-[IN]CORRECT: LLMs Struggle with Discriminating Self-Generated Responses
AAAI 2025
·
Dongwei Jiang
DBLP profile ↗
ORCID search ↗
TurkingBench: A Challenge Benchmark for Web Agents
NAACL 2025
·
Kevin Xu
DBLP profile ↗
ORCID search ↗
Verifiable by Design: Aligning Language Models to Quote from Pre-Training Data
NAACL 2025
·
Jingyu Zhang
DBLP profile ↗
ORCID search ↗
Video-ColBERT: Contextualized Late Interaction for Text-to-Video Retrieval
CVPR 2025
·
Arun V. Reddy
DBLP profile ↗
ORCID search ↗
WorldAPIs: The World Is Worth How Many APIs? A Thought Experiment
AAAI 2025
·
Jiefu Ou
DBLP profile ↗
ORCID search ↗
mFollowIR: A Multilingual Benchmark for Instruction Following in Retrieval
ECIR 2025
·
Orion Weller
DBLP profile ↗
ORCID search ↗
A Survey of Video Datasets for Grounded Event Understanding
CVPR 2024
·
Kate Sanders
DBLP profile ↗
ORCID search ↗
Contrastive Preference Optimization: Pushing the Boundaries of LLM Performance in Machine Translation
ICML 2024
·
Haoran Xu
DBLP profile ↗
ORCID search ↗
Do Androids Know They're Only Dreaming of Electric Sheep?
ACL 2024
·
Sky CH-Wang
DBLP profile ↗
ORCID search ↗
Dodo: Dynamic Contextual Compression for Decoder-only LMs
ACL 2024
·
Guanghui Qin
DBLP profile ↗
ORCID search ↗
Enhancing Systematic Decompositional Natural Language Inference Using Informal Logic
EMNLP 2024
·
Nathaniel Weir
DBLP profile ↗
ORCID search ↗
FAMuS: Frames Across Multiple Sources
NAACL 2024
·
Siddharth Vashishtha
DBLP profile ↗
ORCID search ↗
Grounding Partially-Defined Events in Multimodal Data
EMNLP 2024
·
Kate Sanders
DBLP profile ↗
ORCID search ↗
Interpreting User Requests in the Context of Natural Language Standing Instructions
NAACL 2024
·
Nikita Moghe
DBLP profile ↗
ORCID search ↗
LLM-Rubric: A Multidimensional, Calibrated Approach to Automated Evaluation of Natural Language Texts
ACL 2024
·
Helia Hashemi
DBLP profile ↗
ORCID search ↗
LLMs in the Imaginarium: Tool Learning through Simulated Trial and Error
ACL 2024
·
Boshi Wang
DBLP profile ↗
ORCID search ↗
Language-to-Code Translation with a Single Labeled Example
EMNLP 2024
·
Kaj Bostrom
DBLP profile ↗
ORCID search ↗
Learning to Retrieve Iteratively for In-Context Learning
EMNLP 2024
·
Yunmo Chen
DBLP profile ↗
ORCID search ↗
NELLIE: A Neuro-Symbolic Inference Engine for Grounded, Compositional, and Explainable Reasoning
IJCAI 2024
·
Nathaniel Weir
DBLP profile ↗
ORCID search ↗
Narrowing the Gap between Zero- and Few-shot Machine Translation by Matching Styles
NAACL 2024
·
Weiting Tan
DBLP profile ↗
ORCID search ↗
Natural Language Decomposition and Interpretation of Complex Utterances
IJCAI 2024
·
Harsh Jhamtani
DBLP profile ↗
ORCID search ↗
Ontologically Faithful Generation of Non-Player Character Dialogues
EMNLP 2024
·
Nathaniel Weir
DBLP profile ↗
ORCID search ↗
RORA: Robust Free-Text Rationale Evaluation
ACL 2024
·
Zhengping Jiang
DBLP profile ↗
ORCID search ↗
SemStamp: A Semantic Watermark with Paraphrastic Robustness for Text Generation
NAACL 2024
·
Abe Bohan Hou
DBLP profile ↗
ORCID search ↗
TV-TREES: Multimodal Entailment Trees for Neuro-Symbolic Video Reasoning
EMNLP 2024
·
Kate Sanders
DBLP profile ↗
ORCID search ↗
Zero and Few-shot Semantic Parsing with Ambiguous Inputs
ICLR 2024
·
Elias Stengel-Eskin
DBLP profile ↗
ORCID search ↗