P
PaperPicks
Conferences
Changsheng Xu
50 papers at tracked venues · 46 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID 0000-0001-8343-9665 ↗
Homepage ↗
Venues
ACM MM
×15
AAAI
×8
ICML
×6
NeurIPS
×5
CVPR
×4
ECCV
×3
ICCV
×3
ICLR
×2
WWW
×2
IJCAI
×1
SIGIR
×1
Frequent coauthors
Fan Qi
DBLP profile ↗
ORCID search ↗
×6
Zhenyu Yang
DBLP profile ↗
ORCID search ↗
×4
Feifei Zhang
DBLP profile ↗
ORCID search ↗
×3
Dizhan Xue
DBLP profile ↗
ORCID search ↗
×2
Junyu Gao
DBLP profile ↗
ORCID search ↗
×2
Lu Yu
DBLP profile ↗
ORCID search ↗
×2
Jifei Luo
DBLP profile ↗
ORCID search ↗
×2
Baochen Xiong
DBLP profile ↗
ORCID search ↗
×2
Ming Tao
DBLP profile ↗
ORCID search ↗
×2
Mengyuan Chen
DBLP profile ↗
ORCID search ↗
×2
Linhui Xiao
DBLP profile ↗
ORCID search ↗
×2
Haoliang Zhou
DBLP profile ↗
ORCID search ↗
×1
Papers
Duplex Rewards Optimization for Test-Time Composed Image Retrieval
AAAI 2026
·
Haoliang Zhou
DBLP profile ↗
ORCID search ↗
I2CD: An Invertible Causal Framework for Compositional Zero-Shot Learning via Disentangle-Compose-Disentangle
AAAI 2026
·
Zhaoquan Yuan
DBLP profile ↗
ORCID search ↗
Multi-modal Bipartite Graph Structure Learning with Information Bottleneck for Micro-video Recommendation
WWW 2026
·
Ying He
DBLP profile ↗
ORCID search ↗
SoMe: A Realistic Benchmark for LLM-based Social Media Agents
AAAI 2026
·
Dizhan Xue
DBLP profile ↗
ORCID search ↗
A Large-Scale Dataset for Short-Video Topic Peak Prediction and a Large Heterogeneous Graph Model
ACM MM 2025
·
Shangheng Chen
DBLP profile ↗
ORCID search ↗
Building Embodied EvoAgent: A Brain-inspired Paradigm for Bridging Multimodal Large Models and World Models
ACM MM 2025
·
Junyu Gao
DBLP profile ↗
ORCID search ↗
Customized Condition Controllable Generation for Video Soundtrack
CVPR 2025
·
Fan Qi
DBLP profile ↗
ORCID search ↗
DMC3: Dual-Modal Counterfactual Contrastive Construction for Egocentric Video Question Answering
ACM MM 2025
·
Jiayi Zou
DBLP profile ↗
ORCID search ↗
EgoPrompt: Prompt Learning for Egocentric Action Recognition
ACM MM 2025
·
Huaihai Lyu
DBLP profile ↗
ORCID search ↗
Evidential Knowledge Distillation
ICCV 2025
·
Liangyu Xiang
DBLP profile ↗
ORCID search ↗
FORGET ME: Federated Unlearning for Face Generation Models
ACM MM 2025
·
Fan Qi
DBLP profile ↗
ORCID search ↗
Fine-tuning Bias Neurons for Fair Text-to-Image Generation
ACM MM 2025
·
Fan Qi
DBLP profile ↗
ORCID search ↗
Granular Music Attribute Transformation with Proximal Policy Optimization Adapters for Diffusion Model
ACM MM 2025
·
Kunsheng Ma
DBLP profile ↗
ORCID search ↗
Graph Prompts: Adapting Video Graph for Video Question Answering
IJCAI 2025
·
Yiming Li
DBLP profile ↗
ORCID search ↗
Language Guided Concept Bottleneck Models for Interpretable Continual Learning
CVPR 2025
·
Lu Yu
DBLP profile ↗
ORCID search ↗
LiveStar: Live Streaming Assistant for Real-World Online Video Understanding
NeurIPS 2025
·
Zhenyu Yang
DBLP profile ↗
ORCID search ↗
Locality Preserving Markovian Transition for Instance Retrieval
ICML 2025
·
Jifei Luo
DBLP profile ↗
ORCID search ↗
Look Before You Leap: A GUI-Critic-R1 Model for Pre-Operative Error Diagnosis in GUI Automation
NeurIPS 2025
·
Yuyang Wanyan
DBLP profile ↗
ORCID search ↗
NavMorph: A Self-Evolving World Model for Vision-and-Language Navigation in Continuous Environments
ICCV 2025
·
Xuan Yao
DBLP profile ↗
ORCID search ↗
Overcoming Dual Drift for Continual Long-Tailed Visual Question Answering
ICCV 2025
·
Feifei Zhang
DBLP profile ↗
ORCID search ↗
Pilot: Building the Federated Multimodal Instruction Tuning Framework
AAAI 2025
·
Baochen Xiong
DBLP profile ↗
ORCID search ↗
Pseudo Informative Episode Construction for Few-Shot Class-Incremental Learning
AAAI 2025
·
Chaofan Chen
DBLP profile ↗
ORCID search ↗
Rethinking the Temperature for Federated Heterogeneous Distillation
ICML 2025
·
Fan Qi
DBLP profile ↗
ORCID search ↗
SVBench: A Benchmark with Temporal Multi-Turn Dialogues for Streaming Video Understanding
ICLR 2025
·
Zhenyu Yang
DBLP profile ↗
ORCID search ↗
StreamingCoT: A Dataset for Temporal Dynamics and Multimodal Chain-of-Thought Reasoning in Streaming VideoQA
ACM MM 2025
·
Yuhang Hu
DBLP profile ↗
ORCID search ↗
When Open-Vocabulary Visual Question Answering Meets Causal Adapter: Benchmark and Approach
AAAI 2025
·
Feifei Zhang
DBLP profile ↗
ORCID search ↗
Cluster-Aware Similarity Diffusion for Instance Retrieval
ICML 2024
·
Jifei Luo
DBLP profile ↗
ORCID search ↗
CoIn: A Lightweight and Effective Framework for Story Visualization and Continuation
ACM MM 2024
·
Ming Tao
DBLP profile ↗
ORCID search ↗
Conjugated Semantic Pool Improves OOD Detection with Pre-trained Vision-Language Models
NeurIPS 2024
·
Mengyuan Chen
DBLP profile ↗
ORCID search ↗
Cross-Modal Meta Consensus for Heterogeneous Federated Learning
ACM MM 2024
·
Shuai Li
DBLP profile ↗
ORCID search ↗
Enhancing Storage and Computational Efficiency in Federated Multimodal Learning for Large-Scale Models
ICML 2024
·
Zixin Zhang
DBLP profile ↗
ORCID search ↗
Fast-Slow Test-Time Adaptation for Online Vision-and-Language Navigation
ICML 2024
·
Junyu Gao
DBLP profile ↗
ORCID search ↗
FedVAD: Enhancing Federated Video Anomaly Detection with GPT-Driven Semantic Distillation
ECCV 2024
·
Fan Qi
DBLP profile ↗
ORCID search ↗
Few-Shot Multimodal Explanation for Visual Question Answering
ACM MM 2024
·
Dizhan Xue
DBLP profile ↗
ORCID search ↗
HiVG: Hierarchical Multimodal Fine-grained Modulation for Visual Grounding
ACM MM 2024
·
Linhui Xiao
DBLP profile ↗
ORCID search ↗
LDRE: LLM-based Divergent Reasoning and Ensemble for Zero-Shot Composed Image Retrieval
SIGIR 2024
·
Zhenyu Yang
DBLP profile ↗
ORCID search ↗
Libra: Building Decoupled Vision System on Large Language Models
ICML 2024
·
Yifan Xu
DBLP profile ↗
ORCID search ↗
Modality-Collaborative Test-Time Adaptation for Action Recognition
CVPR 2024
·
Baochen Xiong
DBLP profile ↗
ORCID search ↗
Music Style Transfer with Time-Varying Inversion of Diffusion Models
AAAI 2024
·
Sifei Li
DBLP profile ↗
ORCID search ↗
OneRef: Unified One-tower Expression Grounding and Segmentation with Mask Referring Modeling
NeurIPS 2024
·
Linhui Xiao
DBLP profile ↗
ORCID search ↗
Open-Vocabulary Video Scene Graph Generation via Union-aware Semantic Alignment
ACM MM 2024
·
Ziyue Wu
DBLP profile ↗
ORCID search ↗
Overcoming the Pitfalls of Vision-Language Model for Image-Text Retrieval
ACM MM 2024
·
Feifei Zhang
DBLP profile ↗
ORCID search ↗
R-EDL: Relaxing Nonessential Settings of Evidential Deep Learning
ICLR 2024
·
Mengyuan Chen
DBLP profile ↗
ORCID search ↗
Semantic Editing Increment Benefits Zero-Shot Composed Image Retrieval
ACM MM 2024
·
Zhenyu Yang
DBLP profile ↗
ORCID search ↗
SignGen: End-to-End Sign Language Video Generation with Latent Diffusion
ECCV 2024
·
Fan Qi
DBLP profile ↗
ORCID search ↗
StoryImager: A Unified and Efficient Framework for Coherent Story Visualization and Completion
ECCV 2024
·
Ming Tao
DBLP profile ↗
ORCID search ↗
T3RD: Test-Time Training for Rumor Detection on Social Media
WWW 2024
·
Huaiwen Zhang
DBLP profile ↗
ORCID search ↗
TCP: Textual-Based Class-Aware Prompt Tuning for Visual-Language Model
CVPR 2024
·
Hantao Yao
DBLP profile ↗
ORCID search ↗
Text-Guided Attention is All You Need for Zero-Shot Robustness in Vision-Language Models
NeurIPS 2024
·
Lu Yu
DBLP profile ↗
ORCID search ↗
Three Heads Are Better than One: Complementary Experts for Long-Tailed Semi-supervised Learning
AAAI 2024
·
Chengcheng Ma
DBLP profile ↗
ORCID search ↗