PPaperPicks

Dongyeop Kang

28 papers at tracked venues · 17 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Abstain-R1: Calibrated Abstention and Post-Refusal Clarification via Verifiable RL
  2. Align to Structure: Aligning Large Language Models with Structural Information
  3. Becoming Experienced Judges: Selective Test-Time Learning for Evaluators
  4. Breaking Determinism: Stochastic Modeling for Reliable Off-Policy Evaluation in Ad Auctions
  5. Mary, the Cheeseburger-Eating Vegetarian: Do LLMs Recognize Incoherence in Narratives?
  6. Reasoning Beyond Literal: Cross-style Multimodal Reasoning for Figurative Language Understanding
  7. Scaling Unverifiable Rewards: A Case Study on Visual Insights
  8. ScholaWrite: A Dataset of End-to-End Scholarly Writing
  9. Strong Memory, Weak Control: An Empirical Study of Executive Functioning in LLMs
  10. When Thoughts Meet Facts: Reusable Reasoning for Long-Context LMs
  11. BBScoreV2: Learning Time-Evolution and Latent Alignment from Stochastic Representation
  12. Chain-of-Instructions: Compositional Instruction Tuning on Large Language Models
  13. How LLMs Comprehend Temporal Meaning in Narratives: A Case Study in Cognitive Evaluation of LLMs
  14. ISR-DPO: Aligning Large Multimodal Models for Videos by Iterative Self-Retrospective DPO
  15. Joint Reward and Policy Learning with Demonstrations and Human Feedback Improves Alignment
  16. Learning a High-Quality Robotic Wiping Policy Using Systematic Reward Analysis and Visual-Language Model Based Curriculum
  17. RoSTE: An Efficient Quantization-Aware Supervised Fine-Tuning Approach for Large Language Models
  18. BBScore: A Brownian Bridge Based Metric for Assessing Text Coherence
  19. Benchmarking Cognitive Biases in Large Language Models as Evaluators
  20. Dark Sides: Envisioning, Understanding, and Preventing Harmful Effects of Writing Assistants - The Third Workshop on Intelligent and Interactive Writing Assistants
  21. Dynamic Multi-Reward Weighting for Multi-Style Controllable Generation
  22. How Far Can We Extract Diverse Perspectives from Large Language Models?
  23. II-MMR: Identifying and Improving Multi-modal Multi-hop Reasoning in Visual Question Answering
  24. LearnerVoice: A Dataset of Non-Native English Learners' Spontaneous Speech
  25. Meta-Crafting: Improved Detection of Out-of-Distributed Texts via Crafting Metadata Space (Student Abstract)
  26. Threads of Subtlety: Detecting Machine-Generated Texts Through Discourse Motifs
  27. Tuning Large Multimodal Models for Videos using Reinforcement Learning from AI Feedback
  28. Which Modality should I use - Text, Motif, or Image? : Understanding Graphs with Large Language Models