PPaperPicks

Dirk Hovy

Bocconi University, Milan, Italy

19 papers at tracked venues · 10 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Can Reasoning Help Large Language Models Capture Human Annotator Disagreement?
  2. PATS: Personality-Aware Teaching Strategies with Large Language Model Tutors
  3. Responsible Evaluation of AI for Mental Health
  4. The Pluralistic Moral Gap: Understanding Moral Judgment and Value Differences between Humans and Large Language Models
  5. Beyond Demographics: Fine-tuning Large Language Models to Predict Individuals' Subjective Text Perceptions
  6. Biased Tales: Cultural and Topic Bias in Generating Children's Stories
  7. Co-DETECT: Collaborative Discovery of Edge Cases in Text Classification
  8. Principled Personas: Defining and Measuring the Intended Effects of Persona Prompting on Task Performance
    EMNLP 2025 ·
    Pedro Henrique Luz de Araujo
  9. SafetyPrompts: A Systematic Review of Open Datasets for Evaluating and Improving Large Language Model Safety
  10. The AI Gap: How Socioeconomic Status Affects Language Technology Interactions
  11. "My Answer is C": First-Token Probabilities Do Not Match Text Answers in Instruction-Tuned Language Models
  12. Angry Men, Sad Women: Large Language Models Reflect Gendered Stereotypes in Emotion Attribution
    ACL 2024 ·
    Flor Miriam Plaza del Arco
  13. Classist Tools: Social Class Correlates with Performance in NLP
  14. Compromesso! Italian Many-Shot Jailbreaks undermine the safety of Large Language Models
  15. Divine LLaMAs: Bias, Stereotypes, Stigmatization, and Emotion Representation of Religion in Large Language Models
    EMNLP 2024 ·
    Flor Miriam Plaza del Arco
  16. Narratives at Conflict: Computational Analysis of News Framing in Multilingual Disinformation Campaigns
  17. Political Compass or Spinning Arrow? Towards More Meaningful Evaluations for Values and Opinions in Large Language Models
  18. Twists, Humps, and Pebbles: Multilingual Speech Recognition Models Exhibit Gender Performance Gaps
  19. XSTest: A Test Suite for Identifying Exaggerated Safety Behaviours in Large Language Models