PPaperPicks

Xiaojie Jin

12 papers at tracked venues · 11 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. A Unified Reasoning Framework for Holistic Zero-Shot Video Anomaly Analysis
  2. COCONut-PanCap: Joint Panoptic Segmentation and Grounded Captions for Fine-Grained Understanding and Generation
  3. Flash-Vstream: Efficient Real-Time Understanding for Long Video Streams
  4. VideoWorld: Exploring Knowledge Learning from Unlabeled Videos
  5. COSA: Concatenated Sample Pretrained Vision-Language Foundation Model
  6. Exploring Domain Incremental Video Highlights Detection with the LiveFood Benchmark
  7. MV-Adapter: Multimodal Video Transfer Learning for Video Text Retrieval
    CVPR 2024 · Xiaojie Jin
  8. OSIC: A New One-Stage Image Captioner Coined
  9. PixelLM: Pixel Reasoning with Large Multimodal Model
  10. Stitching Segments and Sentences towards Generalization in Video-Text Pre-training
  11. Video Recognition in Portrait Mode
  12. Vista-llama: Reducing Hallucination in Video Language Models via Equal Distance to Visual Tokens