PPaperPicks

Yiwu Zhong

10 papers at tracked venues · 8 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. DOSE: Data Selection for Multi-Modal LLMs via Off-the-Shelf Models
  2. TextShield-R1: Reinforced Reasoning for Tampered Text Detection
  3. AIM: Adaptive Inference of Multi-Modal LLMs via Token Merging and Pruning
    ICCV 2025 · Yiwu Zhong
  4. Fine-Grained Spatiotemporal Grounding on Egocentric Videos
  5. PAVE: Patching and Adapting Video Large Language Models
  6. Revisiting Tampered Scene Text Detection in the Era of Generative AI
  7. Beyond Embeddings: The Promise of Visual Table in Visual Reasoning
    EMNLP 2024 · Yiwu Zhong
  8. Enhancing Temporal Modeling of Video LLMs via Time Gating
  9. Towards Learning a Generalist Model for Embodied Navigation
  10. Towards Modern Image Manipulation Localization: A Large-Scale Dataset and Novel Methods