PPaperPicks

Qianyi Jiang

5 papers at tracked venues · 5 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. A Token-Level Text Image Foundation Model for Document Understanding
  2. InstructOCR: Instruction Boosting Scene Text Spotting
  3. Marten: Visual Question Answering with Mask Generation for Multi-modal Document Understanding
  4. Multimodal Large Language Models for Text-rich Image Understanding: A Comprehensive Review
  5. ODM: A Text-Image Further Alignment Pre-training Approach for Scene Text Detection and Spotting