P
PaperPicks
Conferences
Huaibo Huang
20 papers at tracked venues · 18 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID 0000-0001-5866-2283 ↗
Homepage ↗
Venues
NeurIPS
×5
CVPR
×4
ACM MM
×3
AAAI
×2
ICCV
×2
ACL
×1
ECCV
×1
EMNLP
×1
ICLR
×1
Frequent coauthors
Qihang Fan
DBLP profile ↗
ORCID search ↗
×4
Yuang Ai
DBLP profile ↗
ORCID search ↗
×4
Xuannan Liu
DBLP profile ↗
ORCID search ↗
×3
Xing Cui
DBLP profile ↗
ORCID search ↗
×2
Xiaotian Han
DBLP profile ↗
ORCID search ↗
×1
Yuguang Zhang
DBLP profile ↗
ORCID search ↗
×1
Tingkai Liu
DBLP profile ↗
ORCID search ↗
×1
Hongbo Wang
DBLP profile ↗
ORCID search ↗
×1
Zi Wang
DBLP profile ↗
ORCID search ↗
×1
Haogeng Liu
DBLP profile ↗
ORCID search ↗
×1
Jin Liu
DBLP profile ↗
ORCID search ↗
×1
Papers
T2Agent: A Tool-augmented Multimodal Misinformation Detection Agent with Monte Carlo Tree Search
AAAI 2026
·
Xing Cui
DBLP profile ↗
ORCID search ↗
Breaking the Low-Rank Dilemma of Linear Attention
CVPR 2025
·
Qihang Fan
DBLP profile ↗
ORCID search ↗
DiCo: Revitalizing ConvNets for Scalable and Efficient Diffusion Modeling
NeurIPS 2025
·
Yuang Ai
DBLP profile ↗
ORCID search ↗
InfiMM-WebMath-40B: Advancing Multimodal Pre-Training for Enhanced Mathematical Reasoning
EMNLP 2025
·
Xiaotian Han
DBLP profile ↗
ORCID search ↗
MMFakeBench: A Mixed-Source Multimodal Misinformation Detection Benchmark for LVLMs
ICLR 2025
·
Xuannan Liu
DBLP profile ↗
ORCID search ↗
Rectifying Magnitude Neglect in Linear Attention
ICCV 2025
·
Qihang Fan
DBLP profile ↗
ORCID search ↗
Semantic Equitable Clustering: A Simple and Effective Strategy for Clustering Vision Tokens
ICCV 2025
·
Qihang Fan
DBLP profile ↗
ORCID search ↗
Video-SafetyBench: A Benchmark for Safety Evaluation of Video LVLMs
NeurIPS 2025
·
Xuannan Liu
DBLP profile ↗
ORCID search ↗
Vision Transformer with Sparse Scan Prior
ACM MM 2025
·
Yuguang Zhang
DBLP profile ↗
ORCID search ↗
DeVAn: Dense Video Annotation for Video-Language Models
ACL 2024
·
Tingkai Liu
DBLP profile ↗
ORCID search ↗
DreamClear: High-Capacity Real-World Image Restoration with Privacy-Safe Dataset Curation
NeurIPS 2024
·
Yuang Ai
DBLP profile ↗
ORCID search ↗
FKA-Owl: Advancing Multimodal Fake News Detection through Knowledge-Augmented LVLMs
ACM MM 2024
·
Xuannan Liu
DBLP profile ↗
ORCID search ↗
Hallo3D: Multi-Modal Hallucination Detection and Mitigation for Consistent 3D Content Generation
NeurIPS 2024
·
Hongbo Wang
DBLP profile ↗
ORCID search ↗
Heterogeneous Test-Time Training for Multi-Modal Person Re-identification
AAAI 2024
·
Zi Wang
DBLP profile ↗
ORCID search ↗
INSTASTYLE: Inversion Noise of a Stylized Image is Secretly a Style Adviser
ECCV 2024
·
Xing Cui
DBLP profile ↗
ORCID search ↗
Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image Restoration
CVPR 2024
·
Yuang Ai
DBLP profile ↗
ORCID search ↗
RMT: Retentive Networks Meet Vision Transformers
CVPR 2024
·
Qihang Fan
DBLP profile ↗
ORCID search ↗
Uncertainty-Aware Source-Free Adaptive Image Super-Resolution with Wavelet Augmentation Transformer
CVPR 2024
·
Yuang Ai
DBLP profile ↗
ORCID search ↗
Visual Anchors Are Strong Information Aggregators For Multimodal Large Language Model
NeurIPS 2024
·
Haogeng Liu
DBLP profile ↗
ORCID search ↗
ZePo: Zero-Shot Portrait Stylization with Faster Sampling
ACM MM 2024
·
Jin Liu
DBLP profile ↗
ORCID search ↗