PPaperPicks

Soyeon Caren Han

24 papers at tracked venues · 14 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. EVOTOOL: Self-Evolving Tool-Use Policy Optimization in LLM Agents via Blame-Aware Mutation and Diversity-Aware Selection
  2. HierCon: Hierarchical Contrastive Attention for Audio Deepfake Detection
  3. MulTiCast: A Multimodal Time Series Forecasting System
  4. MuseKG: An Interactive Knowledge Graph Over Museum Collections
  5. 'No' Matters: Out-of-Distribution Detection in Multimodality Multi-Turn Interactive Dialogue Download PDF
  6. 3M-Game: Multi-Modal Multi-Task Multi-Teacher Learning for Game Event Detection (Student Abstract)
  7. A Training-Free Length Extrapolation Approach for LLMs: Greedy Attention Logit Interpolation
  8. ChuLo: Chunk-Level Key Information Representation for Long Document Understanding
  9. DocDiscNER: Enhanced Document-Level Discontinuous NER via Coordination Ellipses Resolution and Self-Consistency Decoding
  10. Graph-Based Multimodal Contrastive Learning for Chart Question Answering
  11. KIEPrompter: Leveraging Lightweight Models' Predictions for Cost-Effective Key Information Extraction using Vision LLMs
  12. MAGIC-VQA: Multimodal And Grounded Inference with Commonsense Knowledge for Visual Question Answering
  13. MIDAS: Multi-level Intent, Domain, And Slot Knowledge Distillation for Multi-turn NLU
  14. Multimodal Commonsense Knowledge Distillation for Visual Question Answering (Student Abstract)
  15. The 1st International Workshop on Retrieval-driven Generative AI & ScienceON AI Challenge: RDGENAI 2025
  16. TriG-NER: Triplet-Grid Framework for Discontinuous Named Entity Recognition
  17. VRD-IU: Lessons from Visually Rich Document Intelligence and Understanding
  18. 3M-Health: Multimodal Multi-Teacher Knowledge Distillation for Mental Health Detection
  19. 3MVRD: Multimodal Multi-task Multi-teacher Visually-Rich Form Document Understanding
  20. MMVQA: A Comprehensive Dataset for Investigating Multipage Multimodal Information Retrieval in PDF-based Visual Question Answering
  21. MSG-Chart: Multimodal Scene Graph for ChartQA
  22. Multimodal Large Language Models and Tunings: Vision, Language, Sensors, Audio, and Beyond
    ACM MM 2024 · Soyeon Caren Han
  23. PEACH: Pretrained-Embedding Explanation across Contextual and Hierarchical Structure
  24. The Language Model Can Have the Personality: Joint Learning for Personality Enhanced Language Model (Student Abstract)