PPaperPicks

Anna Korhonen

University of Cambridge, Language Technology Laboratory, UK

25 papers at tracked venues · 10 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Can VLMs Predict Future States? Bootstrapping World Models from Inverse Dynamics
  2. Decoupling the Effect of Chain-of-Thought Reasoning: A Human Label Variation Perspective
  3. Dial HEALTHDIAL for Advice: A Multilingual and Multi-Parallel Spoken Dialogue Dataset for Knowledge-Grounded Information Seeking
  4. When Meanings Meet: Investigating the Emergence and Quality of Shared Concept Spaces during Multilingual Language Model Training
  5. A Rose by Any Other Name: LLM-Generated Explanations Are Good Proxies for Human Explanations to Collect Label Distributions on NLI
  6. Cultural Learning-Based Culture Adaptation of Language Models
  7. Explainability and Interpretability of Multilingual Large Language Models: A Survey
  8. Improving Preference Extraction In LLMs By Identifying Latent Knowledge Through Classifying Probes
  9. Iterative Multilingual Spectral Attribute Erasure
  10. Large Language Models are Miscalibrated In-Context Learners
  11. Quantifying Language Disparities in Multilingual Large Language Models
  12. Threading the Needle: Reweaving Chain-of-Thought Reasoning to Explain Human Label Variation
  13. "Seeing the Big through the Small": Can LLMs Approximate Human Judgment Distributions on NLI from a Few Explanations?
  14. Are Large Language Model Temporally Grounded?
  15. CALRec: Contrastive Alignment of Generative LLMs for Sequential Recommendation
  16. DIALIGHT: Lightweight Multilingual Development and Evaluation of Task-Oriented Dialogue Systems with Large Language Models
  17. Fairer Preferences Elicit Improved Human-Aligned Large Language Model Judgments
  18. LongForm: Effective Instruction Tuning with Reverse Instructions
  19. SQATIN: Supervised Instruction Tuning Meets Question Answering for Improved Dialogue NLU
  20. Self-Augmented In-Context Learning for Unsupervised Word Translation
  21. Spectral Editing of Activations for Large Language Model Alignment
  22. SynthEval: Hybrid Behavioral Testing of NLP Models with Synthetic Evaluation
  23. TopViewRS: Vision-Language Models as Top-View Spatial Reasoners
  24. TurkishMMLU: Measuring Massive Multitask Language Understanding in Turkish
  25. Your Prompt Is My Command: On Assessing the Human-Centred Generality of Multimodal Models (Abstract Reprint)