PPaperPicks

Yunlong Tang

15 papers at tracked venues · 14 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Caption Anything in Video: Fine-grained Object-centric Captioning via Spatiotemporal Multimodal Prompting
    AAAI 2026 · Yolo Yunlong Tang
  2. $\pi$-AVAS: Can Physics-Integrated Audio-Visual Modeling Boost Neural Acoustic Synthesis?
  3. CaRDiff: Video Salient Object Ranking Chain of Thought Reasoning for Saliency Prediction with Diffusion
    AAAI 2025 · Yunlong Tang
  4. Empowering LLMs with Pseudo-Untrimmed Videos for Audio-Visual Temporal Understanding
    AAAI 2025 · Yunlong Tang
  5. Generative AI for Cel-Animation: A Survey
    ICCV 2025 · Yunlong Tang
  6. Harnessing the Computation Redundancy in ViTs to Boost Adversarial Transferability
  7. MMIG-Bench: Towards Comprehensive and Explainable Evaluation of Multi-Modal Image Generation Models
  8. MMPerspective: Do MLLMs Understand Perspective? A Comprehensive Benchmark for Perspective Perception, Reasoning, and Robustness
    NeurIPS 2025 · Yunlong Tang
  9. Unveiling Visual Perception in Language Models: An Attention Head Analysis Approach
  10. V2Xum-LLM: Cross-Modal Video Summarization with Temporal Prompt Instruction Tuning
  11. VQualA 2025 Challenge on Engagement Prediction for Short Videos: Methods and Results
  12. VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
    CVPR 2025 · Yunlong Tang
  13. ZeroSep: Separate Anything in Audio with Zero Training
  14. AIM 2024 Challenge on Video Saliency Prediction: Methods and Results
  15. EAGLE: Egocentric AGgregated Language-video Engine