PPaperPicks

Mohammad Taher Pilehvar

University of Cambridge, UK

18 papers at tracked venues · 9 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Synthia: Scalable Grounded Persona Generation from Social Media Data
  2. TruthTrap: A Bilingual Benchmark for Evaluating Factually Correct Yet Misleading Information in Question Answering
  3. Understanding LLM Performance Degradation in Multi-Instance Processing: The Roles of Instance Count and Context Length
  4. Blind Men and the Elephant: Diverse Perspectives on Gender Stereotypes in Benchmark Datasets
  5. Evaluating Cultural Knowledge and Reasoning in LLMs Through Persian Allusions
  6. Findings of the Association for Computational Linguistics, ACL 2025, Vienna, Austria, July 27 - August 1, 2025
  7. LibraGrad: Balancing Gradient Flow for Universally Better Vision Transformer Attributions
  8. Morables: A Benchmark for Assessing Abstract Moral Reasoning in LLMs with Fables
  9. NormXLogit: The Head-on-Top Never Lies
  10. PerCul: A Story-Driven Cultural Evaluation of LLMs in Persian
  11. Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 1: Long Papers), ACL 2025, Vienna, Austria, July 27 - August 1, 2025
  12. Proceedings of the 63rd Annual Meeting of the Association for Computational Linguistics (Volume 2: Short Papers), ACL 2025, Vienna, Austria, July 27 - August 1, 2025
  13. Pun Unintended: LLMs and the Illusion of Humor Understanding
  14. BLEnD: A Benchmark for LLMs on Everyday Knowledge in Diverse Cultures and Languages
  15. RepMatch: Quantifying Cross-Instance Similarities in Representation Space
  16. SkipPLUS: Skip the First Few Layers to Better Explain Vision Transformers
  17. Spanning the Spectrum of Hatred Detection: A Persian Multi-Label Hate Speech Dataset with Annotator Rationales
  18. Stochastic Fine-Tuning of Language Models Using Masked Gradients