PPaperPicks

Kyle Lo

Allen Institute for Artificial Intelligence, Seattle, Washington, USA

23 papers at tracked venues · 16 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. The olmOCR Project: Building Fully Open OCR using VLMs
  2. DrawEduMath: Evaluating Vision Language Models with Expert-Annotated Students' Hand-Drawn Math Images
  3. FlexOLMo: Open Language Models for Flexible Data Use
  4. FollowIR: Evaluating and Teaching Information Retrieval Models to Follow Instructions
  5. Human-AI Collaboration: How AIs Augment Human Teammates
  6. Intent-aware Schema Generation and Refinement for Literature Review Tables
  7. Molmo and PixMo: Open Weights and Open Data for State-of-the-Art Vision-Language Models
  8. OLMoE: Open Mixture-of-Experts Language Models
  9. Organize the Web: Constructing Domains Enhances Pre-Training Data Curation
  10. RouterRetriever: Routing over a Mixture of Expert Embedding Models
  11. SciRIFF: A Resource to Enhance Language Model Instruction-Following over Scientific Literature
  12. Signal and Noise: A Framework for Reducing Uncertainty in Language Model Evaluation
  13. ArxivDIGESTables: Synthesizing Scientific Literature into Tables using Language Models
  14. BooookScore: A systematic exploration of book-length summarization in the era of LLMs
  15. DataComp-LM: In search of the next generation of training sets for language models
  16. Dolma: an Open Corpus of Three Trillion Tokens for Language Model Pretraining Research
  17. InfoLossQA: Characterizing and Recovering Information Loss in Text Simplification
  18. KIWI: A Dataset of Knowledge-Intensive Writing Instructions for Answering Research Questions
  19. Know Your Audience: The benefits and pitfalls of generating plain language summaries beyond the "general" audience
  20. MathFish: Evaluating Language Model Math Reasoning via Grounding in Educational Curricula
  21. OLMo: Accelerating the Science of Language Models
  22. One Thousand and One Pairs: A "novel" challenge for long-context language models
  23. Paloma: A Benchmark for Evaluating Language Model Fit