PPaperPicks

Can Ma

10 papers at tracked venues · 8 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. An Empirical Study on Configuring In-Context Learning Demonstrations for Unleashing MLLMs' Sentimental Perception Capability
  2. Gather and Trace: Rethinking Video TextVQA from an Instance-oriented Perspective
  3. Linguistics-aware Masked Image Modeling for Self-supervised Scene Text Recognition
  4. PACM: Position-Aware Cross-Modality Decoder for Handwritten Mathematical Expression Recognition
  5. SSTAG: Structure-Aware Self-Supervised Learning Method for Text-Attributed Graphs
  6. Track the Answer: Extending TextVQA from Image to Video with Spatio-Temporal Clues
  7. Union Is Strength! Unite the Power of LLMs and MLLMs for Chart Question Answering
  8. Bridging Visual Affective Gap: Borrowing Textual Knowledge by Learning from Noisy Image-Text Pairs
  9. Free your mouse! Command Large Language Models to Generate Code to Format Word Documents
  10. Robust Multimodal Sentiment Analysis of Image-Text Pairs by Distribution-Based Feature Recovery and Fusion