PPaperPicks

Jing Bi

9 papers at tracked venues · 8 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Caption Anything in Video: Fine-grained Object-centric Captioning via Spatiotemporal Multimodal Prompting
  2. Empowering LLMs with Pseudo-Untrimmed Videos for Audio-Visual Temporal Understanding
  3. Generative AI for Cel-Animation: A Survey
  4. MMPerspective: Do MLLMs Understand Perspective? A Comprehensive Benchmark for Perspective Perception, Reasoning, and Robustness
  5. Unveiling Visual Perception in Language Models: An Attention Head Analysis Approach
    CVPR 2025 · Jing Bi
  6. VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
  7. ZeroSep: Separate Anything in Audio with Zero Training
  8. EAGLE: Egocentric AGgregated Language-video Engine
    ACM MM 2024 · Jing Bi
  9. OSCaR: Object State Captioning and State Change Representation