P
PaperPicks
Conferences
Yunlong Tang
15 papers at tracked venues · 14 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID search ↗
Venues
AAAI
×4
NeurIPS
×4
ICCV
×3
CVPR
×2
ACM MM
×1
ECCV
×1
Frequent coauthors
Hang Hua
DBLP profile ↗
ORCID search ↗
×2
Jing Bi
DBLP profile ↗
ORCID search ↗
×2
Susan Liang
DBLP profile ↗
ORCID search ↗
×1
Jiani Liu
DBLP profile ↗
ORCID search ↗
×1
Dasong Li
DBLP profile ↗
ORCID search ↗
×1
Chao Huang
DBLP profile ↗
ORCID search ↗
×1
Andrey Moskalenko
DBLP profile ↗
ORCID search ↗
×1
Papers
Caption Anything in Video: Fine-grained Object-centric Captioning via Spatiotemporal Multimodal Prompting
AAAI 2026
·
Yolo Yunlong Tang
$\pi$-AVAS: Can Physics-Integrated Audio-Visual Modeling Boost Neural Acoustic Synthesis?
ICCV 2025
·
Susan Liang
DBLP profile ↗
ORCID search ↗
CaRDiff: Video Salient Object Ranking Chain of Thought Reasoning for Saliency Prediction with Diffusion
AAAI 2025
·
Yunlong Tang
Empowering LLMs with Pseudo-Untrimmed Videos for Audio-Visual Temporal Understanding
AAAI 2025
·
Yunlong Tang
Generative AI for Cel-Animation: A Survey
ICCV 2025
·
Yunlong Tang
Harnessing the Computation Redundancy in ViTs to Boost Adversarial Transferability
NeurIPS 2025
·
Jiani Liu
DBLP profile ↗
ORCID search ↗
MMIG-Bench: Towards Comprehensive and Explainable Evaluation of Multi-Modal Image Generation Models
NeurIPS 2025
·
Hang Hua
DBLP profile ↗
ORCID search ↗
MMPerspective: Do MLLMs Understand Perspective? A Comprehensive Benchmark for Perspective Perception, Reasoning, and Robustness
NeurIPS 2025
·
Yunlong Tang
Unveiling Visual Perception in Language Models: An Attention Head Analysis Approach
CVPR 2025
·
Jing Bi
DBLP profile ↗
ORCID search ↗
V2Xum-LLM: Cross-Modal Video Summarization with Temporal Prompt Instruction Tuning
AAAI 2025
·
Hang Hua
DBLP profile ↗
ORCID search ↗
VQualA 2025 Challenge on Engagement Prediction for Short Videos: Methods and Results
ICCV 2025
·
Dasong Li
DBLP profile ↗
ORCID search ↗
VidComposition: Can MLLMs Analyze Compositions in Compiled Videos?
CVPR 2025
·
Yunlong Tang
ZeroSep: Separate Anything in Audio with Zero Training
NeurIPS 2025
·
Chao Huang
DBLP profile ↗
ORCID search ↗
AIM 2024 Challenge on Video Saliency Prediction: Methods and Results
ECCV 2024
·
Andrey Moskalenko
DBLP profile ↗
ORCID search ↗
EAGLE: Egocentric AGgregated Language-video Engine
ACM MM 2024
·
Jing Bi
DBLP profile ↗
ORCID search ↗