PPaperPicks

Yejin Choi

Stanford University, Department of Computer Science and Institute for Human-Centered AI (HAI), Stanford, CA, USA

79 papers at tracked venues · 57 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Nemotron-CrossThink: Scaling Self-Learning beyond Math Reasoning
  2. When One LLM Drools, Multi-LLM Collaboration Rules
  3. AI Debate Aids Assessment of Controversial Claims
  4. AI as Humanity's Salieri: Quantifying Linguistic Creativity of Language Models via Systematic Attribution of Machine Text against Web Text
  5. ALPACA AGAINST VICUNA: Using LLMs to Uncover Memorization of LLMs
  6. Artificial Hivemind: The Open-Ended Homogeneity of Language Models (and Beyond)
  7. BLIP-3: A Family of Open Large Multimodal Models
  8. Benchmarking Vision Language Model Unlearning via Fictitious Facial Identity Dataset
  9. Bias in Gender Bias Benchmarks: How Spurious Features Distort Evaluation
  10. Biased LLMs can Influence Political Decision-Making
  11. Broken Tokens? Your Language Model can Secretly Handle Non-Canonical Tokenizations
  12. Can Language Models Reason about Individualistic Human Values and Preferences?
  13. CertainlyUncertain: A Benchmark and Metric for Multimodal Epistemic and Aleatoric Awareness
  14. CulturalBench: A Robust, Diverse and Challenging Benchmark for Measuring LMs' Cultural Knowledge Through Human-AI Red-Teaming
  15. DailyDilemmas: Revealing Value Preferences of LLMs with Quandaries of Daily Life
  16. Diverging Preferences: When do Annotators Disagree and do Models Know?
  17. Explore Theory of Mind: program-guided adversarial data generation for theory of mind reasoning
  18. Guardrails and Security for LLMs: Safe, Secure and Controllable Steering of LLM Applications
  19. HALoGEN: Fantastic LLM Hallucinations and Where to Find Them
  20. Infini-gram mini: Exact n-gram Search at the Internet Scale with FM-Index
  21. Information-Guided Identification of Training Data Imprint in (Proprietary) Large Language Models
  22. L3GO: Language Agents with Chain-of-3D-Thoughts for Generating Unconventional Objects
  23. LOTUS: A Leaderboard for Detailed Image Captioning from Quality to Societal Bias and User Preferences
  24. Language Model Alignment in Multilingual Trolley Problems
  25. Machine Unlearning Doesn't Do What You Think: Lessons for Generative AI Policy and Research
  26. Magpie: Alignment Data Synthesis from Scratch by Prompting Aligned LLMs with Nothing
  27. Making VLMs More Robot-Friendly: Self-Critical Distillation of Low-Level Procedural Reasoning
  28. Model Swarms: Collaborative Search to Adapt LLM Experts via Swarm Intelligence
  29. OLMoTrace: Tracing Language Model Outputs Back to Trillions of Training Tokens
  30. One-Minute Video Generation with Test-Time Training
  31. Position: Political Neutrality in AI Is Impossible - But Here Is How to Approximate It
  32. Prismatic Synthesis: Gradient-based Data Diversification Boosts Generalization in LLM Reasoning
  33. ProRL: Prolonged Reinforcement Learning Expands Reasoning Boundaries in Large Language Models
  34. RewardBench: Evaluating Reward Models for Language Modeling
  35. SafetyAnalyst: Interpretable, Transparent, and Steerable Safety Moderation for AI Behavior
  36. Socratic-MCTS: Test-Time Visual Reasoning by Asking the Right Questions
  37. Synthetic Visual Genome
  38. The BiGGen Bench: A Principled Benchmark for Fine-grained Evaluation of Language Models with Language Models
  39. Trust or Escalate: LLM Judges with Provable Guarantees for Human Agreement
  40. VAGEN: Reinforcing World Model Reasoning for Multi-Turn VLM Agents
  41. Why and How LLMs Hallucinate: Connecting the Dots with Subsequence Associations
  42. WildBench: Benchmarking LLMs with Challenging Tasks from Real Users in the Wild
  43. ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning
  44. ActionAtlas: A VideoQA Benchmark for Domain-specialized Action Recognition
  45. Agent Lumos: Unified and Modular Training for Open-Source Language Agents
  46. Can LLMs Keep a Secret? Testing Privacy Implications of Language Models via Contextual Integrity Theory
  47. Can LLMs Reason with Rules? Logic Scaffolding for Stress-Testing and Improving LLMs
  48. CopyBench: Measuring Literal and Non-Literal Reproduction of Copyright-Protected Text in Language Model Generation
  49. Data Mixture Inference Attack: BPE Tokenizers Reveal Training Data Compositions
  50. How to Train Your Fact Verifier: Knowledge Transfer with Multimodal Open Models
  51. Impossible Distillation for Paraphrasing and Summarization: How to Make High-quality Lemonade out of Small, Low-quality Model
  52. In Search of the Long-Tail: Systematic Generation of Long-Tail Inferential Knowledge via Logical Rule Guided Search
  53. JAMDEC: Unsupervised Authorship Obfuscation using Constrained Decoding over Small Language Models
  54. MINT-1T: Scaling Open-Source Multimodal Data by 10x: A Multimodal Dataset with One Trillion Tokens
  55. MacGyver: Are Large Language Models Creative Problem Solvers?
  56. Modular Pluralism: Pluralistic Alignment via Multi-LLM Collaboration
  57. NeuroComparatives: Neuro-Symbolic Distillation of Comparative Knowledge
  58. Perceptions to Beliefs: Exploring Precursory Inferences for Theory of Mind in Large Language Models
  59. Phenomenal Yet Puzzling: Testing Inductive Reasoning Capabilities of Language Models with Hypothesis Refinement
  60. PlaSma: Procedural Knowledge Models for Language-based Planning and Re-Planning
  61. Position: A Roadmap to Pluralistic Alignment
  62. Prometheus: Inducing Fine-Grained Evaluation Capability in Language Models
  63. Quantifying Language Models' Sensitivity to Spurious Features in Prompt Design or: How I learned to start worrying about prompt formatting
  64. Selective "Selective Prediction": Reducing Unnecessary Abstention in Vision-Language Reasoning
  65. Structured Chemistry Reasoning with Large Language Models
  66. StyleRemix: Interpretable Authorship Obfuscation via Distillation and Perturbation of Style Elements
  67. Symbolic Working Memory Enhances Language Models for Complex Rule Application
  68. Tailoring Self-Rationalizers with Multi-Reward Distillation
  69. The Art of Saying No: Contextual Noncompliance in Language Models
  70. The Generative AI Paradox: "What It Can Create, It May Not Understand"
  71. The Unlocking Spell on Base LLMs: Rethinking Alignment via In-Context Learning
  72. UNcommonsense Reasoning: Abductive Reasoning about Uncommon Situations
  73. Unpacking DPO and PPO: Disentangling Best Practices for Learning from Preference Feedback
  74. Value Kaleidoscope: Engaging AI with Pluralistic Human Values, Rights, and Duties
  75. WildChat: 1M ChatGPT Interaction Logs in the Wild
  76. WildGuard: Open One-stop Moderation Tools for Safety Risks, Jailbreaks, and Refusals of LLMs
  77. WildTeaming at Scale: From In-the-Wild Jailbreaks to (Adversarially) Safer Language Models
  78. WildVis: Open Source Visualizer for Million-Scale Chat Logs in the Wild
  79. WildVision: Evaluating Vision-Language Models in the Wild with Human Preferences