PPaperPicks

Mrinmaya Sachan

58 papers at tracked venues · 30 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Bridging Instead of Replacing Online Coding Communities with AI through Community-Enriched Chatbot Designs CSCW008
  2. Can Reasoning Help Large Language Models Capture Human Annotator Disagreement?
  3. Early-Exit and Instant Confidence Translation Quality Estimation
  4. Efficient Test-Time Scaling of Multi-Step Reasoning by Probing Internal States of Large Language Models
  5. Faithfulness-Aware Uncertainty Quantification for Fact-Checking the Output of Retrieval-Augmented Generation
  6. PATS: Personality-Aware Teaching Strategies with Large Language Model Tutors
  7. PaperMentor: A Human-Centered Multi-Agent Writing Tutor for AI Research Papers in Overleaf
  8. Tackling the Root of Misinformation by Teaching Laypeople about Logical Fallacies via Socratic Questioning and Critical Argumentation
  9. Test of Time: Rethinking Temporal Signal of Benchmark Contamination
  10. ThinkBooster: A Unified Framework for Seamless Test-Time Scaling of LLM Reasoning
  11. Uncovering Hidden Correctness in LLM Causal Reasoning via Symbolic Verification
  12. A Head to Predict and a Head to Question: Pre-trained Uncertainty Quantification Heads for Hallucination Detection in LLM Outputs
  13. AI-Assisted Human Evaluation of Machine Translation
  14. Are Language Models Efficient Reasoners? A Perspective from Logic Programming
  15. Calibrating Large Language Models with Sample Consistency
  16. Can Vision-Language Models Solve Visual Math Equations?
    EMNLP 2025 ·
    Monjoy Narayan Choudhury
  17. Co-DETECT: Collaborative Discovery of Edge Cases in Text Classification
  18. DIRAS: Efficient LLM Annotation of Document Relevance for Retrieval Augmented Generation
  19. Dense SAE Latents Are Features, Not Bugs
  20. Do Vision-Language Models Really Understand Visual Language?
  21. From Problem-Solving to Teaching Problem-Solving: Aligning LLMs with Pedagogy using Reinforcement Learning
  22. GPT-4 as a Homework Tutor Can Improve Student Engagement and Learning Outcomes
  23. Generating Pedagogically Meaningful Visuals for Math Word Problems: A New Benchmark and Analysis of Text-to-Image Models
  24. Grammar Control in Dialogue Response Generation for Language Learning Chatbots
  25. Improving Large Language Model Safety with Contrastive Representation Learning
  26. Investigating the Zone of Proximal Development of Language Models for In-Context Learning
  27. Language Model Alignment in Multilingual Trolley Problems
  28. MathGAP: Out-of-Distribution Evaluation on Problems with Arbitrarily Complex Proofs
  29. MathTutorBench: A Benchmark for Measuring Open-ended Pedagogical Capabilities of LLM Tutors
  30. Personalized Exercise Recommendation with Semantically-Grounded Knowledge Tracing
  31. Pointwise Mutual Information as a Performance Gauge for Retrieval-Augmented Generation
  32. Probing for Arithmetic Errors in Language Models
  33. SIKeD: Self-guided Iterative Knowledge Distillation for Mathematical Reasoning
  34. SeePhys: Does Seeing Help Thinking? - Benchmarking Vision-Based Physics Reasoning
  35. The Medium Is Not the Message: Deconfounding Document Embeddings via Linear Concept Erasure
  36. Towards the Pedagogical Steering of Large Language Models for Tutoring: A Case Study with Modeling Productive Failure
  37. A Transformer with Stack Attention
  38. AFaCTA: Assisting the Annotation of Factual Claim Detection with Reliable LLM Annotators
  39. Book2Dial: Generating Teacher Student Interactions from Textbooks for Cost-Effective Development of Educational Chatbots
  40. Can Large Language Models Infer Causation from Correlation?
  41. CausalCite: A Causal Formulation of Paper Citations
  42. Competition of Mechanisms: Tracing How Language Models Handle Facts and Counterfactuals
  43. Confidence Regulation Neurons in Language Models
  44. Cooperate or Collapse: Emergence of Sustainable Cooperation in a Society of LLM Agents
  45. Do LLMs Think Fast and Slow? A Causal Study on Sentiment Analysis
  46. Do Language Models Exhibit the Same Cognitive Biases in Problem Solving as Human Learners?
  47. Efficiently Computing Susceptibility to Context in Language Models
  48. Elastic Weight Removal for Faithful and Abstractive Dialogue Generation
  49. How to Engage your Readers? Generating Guiding Questions to Promote Active Reading
  50. Implicit Personalization in Language Models: A Systematic Study
  51. On Affine Homotopy between Language Encoders
  52. RELIC: Investigating Large Language Model Responses using Self-Consistency
  53. RETRO-LI: Small-Scale Retrieval Augmented Generation Supporting Noisy Similarity Searches and Domain Shift Generalization
  54. Slicing, Chatting, and Refining: A Concept-Based Approach for Machine Learning Model Validation with ConceptSlicer
  55. Stepwise Verification and Remediation of Student Reasoning Errors with Large Language Model Tutors
  56. The ART of LLM Refinement: Ask, Refine, and Trust
  57. Towards Aligning Language Models with Textual Feedback
  58. What Do Language Models Learn in Context? The Structured Task Hypothesis