P
PaperPicks
Conferences
Zuxuan Wu
43 papers at tracked venues · 37 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID 0000-0002-8689-5807 ↗
Venues
CVPR
×9
ICCV
×9
NeurIPS
×9
AAAI
×6
ECCV
×4
ACM MM
×2
ACL
×1
EMNLP
×1
ICLR
×1
IJCAI
×1
Frequent coauthors
Haoran Chen
DBLP profile ↗
ORCID search ↗
×3
Hui Zhang
DBLP profile ↗
ORCID search ↗
×3
Shuyuan Tu
DBLP profile ↗
ORCID search ↗
×3
Junke Wang
DBLP profile ↗
ORCID search ↗
×3
Zhen Xing
DBLP profile ↗
ORCID search ↗
×2
Zhenxin Li
DBLP profile ↗
ORCID search ↗
×2
Wujian Peng
DBLP profile ↗
ORCID search ↗
×2
Rui Tian
DBLP profile ↗
ORCID search ↗
×2
Lingchen Meng
DBLP profile ↗
ORCID search ↗
×2
Wenhao Yao
DBLP profile ↗
ORCID search ↗
×1
Sicheng Xie
DBLP profile ↗
ORCID search ↗
×1
Zhiheng Xi
DBLP profile ↗
ORCID search ↗
×1
Papers
DriveSuprim: Towards Precise Trajectory Selection for End-to-End Planning
AAAI 2026
·
Wenhao Yao
DBLP profile ↗
ORCID search ↗
Human2Robot: Learning Robot Actions from Paired Human-Robot Videos
AAAI 2026
·
Sicheng Xie
DBLP profile ↗
ORCID search ↗
Achieving More with Less: Additive Prompt Tuning for Rehearsal-Free Class-Incremental Learning
ICCV 2025
·
Haoran Chen
DBLP profile ↗
ORCID search ↗
AdaDiff: Adaptive Step Selection for Fast Diffusion Models
AAAI 2025
·
Hui Zhang
DBLP profile ↗
ORCID search ↗
Adaptive Retention & Correction: Test-Time Training for Continual Learning
ICLR 2025
·
Haoran Chen
DBLP profile ↗
ORCID search ↗
AgentGym: Evaluating and Training Large Language Model-based Agents across Diverse Environments
ACL 2025
·
Zhiheng Xi
DBLP profile ↗
ORCID search ↗
Aid: Adapting Image2video Diffusion Models for Instruction-Guided Video Prediction
ICCV 2025
·
Zhen Xing
DBLP profile ↗
ORCID search ↗
BlockDance: Reuse Structurally Similar Spatio-Temporal Features to Accelerate Diffusion Transformers
CVPR 2025
·
Hui Zhang
DBLP profile ↗
ORCID search ↗
Comprehensive Multi-Modal Prototypes Are Simple and Effective Classifiers for Vast-Vocabulary Object Detection
AAAI 2025
·
Yitong Chen
DBLP profile ↗
ORCID search ↗
CreatiLayout: Siamese Multimodal Diffusion Transformer for Creative Layout-to-Image Generation
ICCV 2025
·
Hui Zhang
DBLP profile ↗
ORCID search ↗
EDEN: Enhanced Diffusion for High-quality Large-motion Video Frame Interpolation
CVPR 2025
·
Zihao Zhang
DBLP profile ↗
ORCID search ↗
FNIN: A Fourier Neural Operator-based Numerical Integration Network for Surface-from-gradients
AAAI 2025
·
Jiaqi Leng
DBLP profile ↗
ORCID search ↗
FOCUS: Towards Universal Foreground Segmentation
AAAI 2025
·
Zuyao You
DBLP profile ↗
ORCID search ↗
ForgerySleuth: Empowering Multimodal Large Language Models for Image Manipulation Detection
NeurIPS 2025
·
Zhihao Sun
DBLP profile ↗
ORCID search ↗
Hydra-NeXt: Robust Closed-Loop Driving with Open-Loop Training
ICCV 2025
·
Zhenxin Li
DBLP profile ↗
ORCID search ↗
INST-IT: Boosting Instance Understanding via Explicit Visual Prompt Instruction Tuning
NeurIPS 2025
·
Wujian Peng
DBLP profile ↗
ORCID search ↗
MagicMotion: Controllable Video Generation with Dense-to-Sparse Trajectory Guidance
ICCV 2025
·
Quanhao Li
DBLP profile ↗
ORCID search ↗
MotionFollower: Editing Video Motion via Score-Guided Diffusion
ICCV 2025
·
Shuyuan Tu
DBLP profile ↗
ORCID search ↗
OmniGen-AR: AutoRegressive Any-to-Image Generation
NeurIPS 2025
·
Junke Wang
DBLP profile ↗
ORCID search ↗
ProLongVid: A Simple but Strong Baseline for Long-context Video Instruction Tuning
EMNLP 2025
·
Rui Wang
DBLP profile ↗
ORCID search ↗
REDUCIO! Generating 1K Video Within 16 Seconds Using Extremely Compressed Motion Latents
ICCV 2025
·
Rui Tian
DBLP profile ↗
ORCID search ↗
Rethinking Discrete Tokens: Treating Them as Conditions for Continuous Autoregressive Image Synthesis
ICCV 2025
·
Peng Zheng
DBLP profile ↗
ORCID search ↗
Seg2Any: Open-set Segmentation-Mask-to-Image Generation with Precise Shape and Semantic Control
NeurIPS 2025
·
Danfeng Li
DBLP profile ↗
ORCID search ↗
StableAnimator: High-Quality Identity-Preserving Human Image Animation
CVPR 2025
·
Shuyuan Tu
DBLP profile ↗
ORCID search ↗
UniGen: Enhanced Training & Test-Time Strategies for Unified Multimodal Understanding and Generation
NeurIPS 2025
·
Rui Tian
DBLP profile ↗
ORCID search ↗
VLABench: A Large-Scale Benchmark for Language-Conditioned Robotics Manipulation with Long-Horizon Reasoning Tasks
ICCV 2025
·
Shiduo Zhang
DBLP profile ↗
ORCID search ↗
Aligning Vision Models with Human Aesthetics in Retrieval: Benchmarks and Algorithms
NeurIPS 2024
·
Miaosen Zhang
DBLP profile ↗
ORCID search ↗
BEVNeXt: Reviving Dense BEV Frameworks for 3D Object Detection
CVPR 2024
·
Zhenxin Li
DBLP profile ↗
ORCID search ↗
DeepStack: Deeply Stacking Visual Tokens is Surprisingly Simple and Effective for LMMs
NeurIPS 2024
·
Lingchen Meng
DBLP profile ↗
ORCID search ↗
DreamMesh: Jointly Manipulating and Texturing Triangle Meshes for Text-to-3D Generation
ECCV 2024
·
Haibo Yang
DBLP profile ↗
ORCID search ↗
Fuse Your Latents: Video Editing with Multi-source Latent Diffusion Models
ACM MM 2024
·
Tianyi Lu
DBLP profile ↗
ORCID search ↗
GenRec: Unifying Video Generation and Recognition with Diffusion Models
NeurIPS 2024
·
Zejia Weng
DBLP profile ↗
ORCID search ↗
Learning to Rank Patches for Unbiased Image Redundancy Reduction
CVPR 2024
·
Yang Luo
DBLP profile ↗
ORCID search ↗
MagDiff: Multi-alignment Diffusion for High-Fidelity Video Generation and Editing
ECCV 2024
·
Haoyu Zhao
DBLP profile ↗
ORCID search ↗
ModelLock: Locking Your Model With a Spell
ACM MM 2024
·
Yifeng Gao
DBLP profile ↗
ORCID search ↗
MotionEditor: Editing Video Motion via Content-Aware Diffusion
CVPR 2024
·
Shuyuan Tu
DBLP profile ↗
ORCID search ↗
OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation
NeurIPS 2024
·
Junke Wang
DBLP profile ↗
ORCID search ↗
OmniViD: A Generative Framework for Universal Video Understanding
CVPR 2024
·
Junke Wang
DBLP profile ↗
ORCID search ↗
PromptFusion: Decoupling Stability and Plasticity for Continual Learning
ECCV 2024
·
Haoran Chen
DBLP profile ↗
ORCID search ↗
SegIC: Unleashing the Emergent Correspondence for In-Context Segmentation
ECCV 2024
·
Lingchen Meng
DBLP profile ↗
ORCID search ↗
SimDA: Simple Diffusion Adapter for Efficient Video Generation
CVPR 2024
·
Zhen Xing
DBLP profile ↗
ORCID search ↗
Synthesize, Diagnose, and Optimize: Towards Fine-Grained Vision-Language Understanding
CVPR 2024
·
Wujian Peng
DBLP profile ↗
ORCID search ↗
Zero-shot High-fidelity and Pose-controllable Character Animation
IJCAI 2024
·
Bingwen Zhu
DBLP profile ↗
ORCID search ↗