PPaperPicks

Kunchang Li

13 papers at tracked venues · 11 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. Bootstrapping Language-Guided Navigation Learning with Self-Refining Data Flywheel
  2. Make Your Training Flexible: Towards Deployment-Efficient Video Models
  3. Muses: 3D-Controllable Image Generation via Multi-Modal Agent Collaboration
  4. Task Preference Optimization: Improving Multimodal Large Language Models with Vision Task Alignment
  5. TimeStep Master: Asymmetrical Mixture of Timestep LoRA Experts for Versatile and Efficient Diffusion Models in Vision
  6. TimeSuite: Improving MLLMs for Long Video Understanding via Grounded Tuning
  7. V-Stylist: Video Stylization via Collaboration and Reflection of MLLM Agents
  8. InternVid: A Large-scale Video-Text Dataset for Multimodal Understanding and Generation
  9. InternVideo2: Scaling Foundation Models for Multimodal Video Understanding
  10. MVBench: A Comprehensive Multi-modal Video Understanding Benchmark
    CVPR 2024 · Kunchang Li
  11. TransAgent: Transfer Vision-Language Foundation Models with Heterogeneous Agent Collaboration
  12. VideoMamba: State Space Model for Efficient Video Understanding
    ECCV 2024 · Kunchang Li
  13. Vlogger: Make Your Dream A Vlog