PPaperPicks

Shangzhe Di

6 papers at tracked venues · 6 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. Enhancing Video-LLM Reasoning via Agent-of-Thoughts Distillation
  2. Grounded Multi-Hop VideoQA in Long-Form Egocentric Videos
  3. Learning Streaming Video Representation via Multitask Training
  4. Streaming Video Question-Answering with In-context Video KV-Cache Retrieval
    ICLR 2025 · Shangzhe Di
  5. Universal Video Temporal Grounding with Generative Multi-modal Large Language Models
  6. Grounded Question-Answering in Long Egocentric Videos
    CVPR 2024 · Shangzhe Di