P
PaperPicks
Conferences
Shayne Longpre
21 papers at tracked venues · 18 at CORE A* · active 2024–2025
DBLP profile ↗
ORCID search ↗
Venues
NeurIPS
×6
ICML
×5
ICLR
×4
ACL
×2
NAACL
×2
AAAI
×1
EMNLP
×1
Frequent coauthors
Seungone Kim
DBLP profile ↗
ORCID search ↗
×3
Shivalika Singh
DBLP profile ↗
ORCID search ↗
×2
Yuxuan Zhu
DBLP profile ↗
ORCID search ↗
×1
Weijia Shi
DBLP profile ↗
ORCID search ↗
×1
Nikhil Kandpal
DBLP profile ↗
ORCID search ↗
×1
Sean McGregor
DBLP profile ↗
ORCID search ↗
×1
Yiwei Wu
DBLP profile ↗
ORCID search ↗
×1
Ahmet Üstün
DBLP profile ↗
ORCID search ↗
×1
Sheng Shen
DBLP profile ↗
ORCID search ↗
×1
Niklas Muennighoff
DBLP profile ↗
ORCID search ↗
×1
Riley Simmons-Edler
DBLP profile ↗
ORCID search ↗
×1
Sayash Kapoor
DBLP profile ↗
ORCID search ↗
×1
Papers
Bridging the Data Provenance Gap Across Text, Speech, and Video
ICLR 2025
·
Shayne Longpre
Establishing Best Practices in Building Rigorous Agentic Benchmarks
NeurIPS 2025
·
Yuxuan Zhu
DBLP profile ↗
ORCID search ↗
FlexOLMo: Open Language Models for Flexible Data Use
NeurIPS 2025
·
Weijia Shi
DBLP profile ↗
ORCID search ↗
Global MMLU: Understanding and Addressing Cultural and Linguistic Biases in Multilingual Evaluation
ACL 2025
·
Shivalika Singh
DBLP profile ↗
ORCID search ↗
Position: In-House Evaluation Is Not Enough. Towards Robust Third-Party Evaluation and Flaw Disclosure for General-Purpose AI
ICML 2025
·
Shayne Longpre
The BiGGen Bench: A Principled Benchmark for Fine-grained Evaluation of Language Models with Language Models
NAACL 2025
·
Seungone Kim
DBLP profile ↗
ORCID search ↗
The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text
NeurIPS 2025
·
Nikhil Kandpal
DBLP profile ↗
ORCID search ↗
The Leaderboard Illusion
NeurIPS 2025
·
Shivalika Singh
DBLP profile ↗
ORCID search ↗
To Err Is AI: A Case Study Informing LLM Flaw Reporting Practices
AAAI 2025
·
Sean McGregor
DBLP profile ↗
ORCID search ↗
A Pretrainer's Guide to Training Data: Measuring the Effects of Data Age, Domain Coverage, Quality, & Toxicity
NAACL 2024
·
Shayne Longpre
A Systematic Review of NeurIPS Dataset Management Practices
NeurIPS 2024
·
Yiwei Wu
DBLP profile ↗
ORCID search ↗
Aya Model: An Instruction Finetuned Open-Access Multilingual Language Model
ACL 2024
·
Ahmet Üstün
DBLP profile ↗
ORCID search ↗
Consent in Crisis: The Rapid Decline of the AI Data Commons
NeurIPS 2024
·
Shayne Longpre
Mixture-of-Experts Meets Instruction Tuning: A Winning Combination for Large Language Models
ICLR 2024
·
Sheng Shen
DBLP profile ↗
ORCID search ↗
OctoPack: Instruction Tuning Code Large Language Models
ICLR 2024
·
Niklas Muennighoff
DBLP profile ↗
ORCID search ↗
Position: A Safe Harbor for AI Evaluation and Red Teaming
ICML 2024
·
Shayne Longpre
Position: AI-Powered Autonomous Weapons Risk Geopolitical Instability and Threaten AI Research
ICML 2024
·
Riley Simmons-Edler
DBLP profile ↗
ORCID search ↗
Position: Data Authenticity, Consent, & Provenance for AI are all broken: what will it take to fix them?
ICML 2024
·
Shayne Longpre
Position: On the Societal Impact of Open Foundation Models
ICML 2024
·
Sayash Kapoor
DBLP profile ↗
ORCID search ↗
Prometheus 2: An Open Source Language Model Specialized in Evaluating Other Language Models
EMNLP 2024
·
Seungone Kim
DBLP profile ↗
ORCID search ↗
Prometheus: Inducing Fine-Grained Evaluation Capability in Language Models
ICLR 2024
·
Seungone Kim
DBLP profile ↗
ORCID search ↗