PPaperPicks

Zhengyang Liang

8 papers at tracked venues · 8 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. Any Information Is Just Worth One Single Screenshot: Unifying Search With Visualized Information Retrieval
  2. Dynamic Self-adaptive Multiscale Distillation from Pre-trained Multimodal Large Model for Efficient Cross-modal Retrieval
    ACM MM 2025 · Zhengyang Liang
  3. MLVU: Benchmarking Multi-task Long Video Understanding
  4. MomentSeeker: A Task-Oriented Benchmark For Long-Video Moment Retrieval
  5. Unveiling the Ignorance of MLLMs: Seeing Clearly, Answering Incorrectly
  6. Video-XL: Extra-Long Vision Language Model for Hour-Scale Video Understanding
  7. AnimateDiff: Animate Your Personalized Text-to-Image Diffusion Models without Specific Tuning
  8. Self-Supervised Multi-Modal Knowledge Graph Contrastive Hashing for Cross-Modal Search