P
PaperPicks
Conferences
Jianhua Han
21 papers at tracked venues · 15 at CORE A* · active 2024–2025
DBLP profile ↗
ORCID search ↗
Venues
ECCV
×5
NeurIPS
×5
CVPR
×4
ICLR
×3
AAAI
×1
ACL
×1
ICCV
×1
WACV
×1
Frequent coauthors
Ming Nie
DBLP profile ↗
ORCID search ↗
×3
Youpeng Wen
DBLP profile ↗
ORCID search ↗
×2
Runhui Huang
DBLP profile ↗
ORCID search ↗
×2
Kai Chen
DBLP profile ↗
ORCID search ↗
×1
Jiahui Gao
DBLP profile ↗
ORCID search ↗
×1
Chunwei Wang
DBLP profile ↗
ORCID search ↗
×1
Kun Xiang
DBLP profile ↗
ORCID search ↗
×1
Qingping Zheng
DBLP profile ↗
ORCID search ↗
×1
Xiwen Liang
DBLP profile ↗
ORCID search ↗
×1
Lewei Yao
DBLP profile ↗
ORCID search ↗
×1
Kai Chen
DBLP profile ↗
ORCID search ↗
×1
Xinpeng Ding
DBLP profile ↗
ORCID search ↗
×1
Papers
DisCo: Discovering Common Affordance from Large Models for Actionable Part Perception
WACV 2025
·
Youpeng Wen
DBLP profile ↗
ORCID search ↗
EMOVA: Empowering Language Models to See, Hear and Speak with Vivid Emotions
CVPR 2025
·
Kai Chen
DBLP profile ↗
ORCID search ↗
G-LLaVA: Solving Geometric Problem with Multi-Modal Large Language Model
ICLR 2025
·
Jiahui Gao
DBLP profile ↗
ORCID search ↗
HiRes-LLaVA: Restoring Fragmentation Input in High-Resolution Large Vision-Language Models
CVPR 2025
·
Runhui Huang
DBLP profile ↗
ORCID search ↗
ILLUME: Illuminating Your LLMs to See, Draw, and Self-Enhance
ICCV 2025
·
Chunwei Wang
DBLP profile ↗
ORCID search ↗
SeePhys: Does Seeing Help Thinking? - Benchmarking Vision-Based Physics Reasoning
NeurIPS 2025
·
Kun Xiang
DBLP profile ↗
ORCID search ↗
Towards Unified Multimodal Interleaved Generation via Group Relative Policy Optimization
NeurIPS 2025
·
Ming Nie
DBLP profile ↗
ORCID search ↗
Any-Size-Diffusion: Toward Efficient Text-Driven Synthesis for Any-Size HD Images
AAAI 2024
·
Qingping Zheng
DBLP profile ↗
ORCID search ↗
CorNav: Autonomous Agent with Self-Corrected Planning for Zero-Shot Vision-and-Language Navigation
ACL 2024
·
Xiwen Liang
DBLP profile ↗
ORCID search ↗
DetCLIPv3: Towards Versatile Generative Open-Vocabulary Object Detection
CVPR 2024
·
Lewei Yao
DBLP profile ↗
ORCID search ↗
Gaining Wisdom from Setbacks: Aligning Large Language Models via Mistake Analysis
ICLR 2024
·
Kai Chen
DBLP profile ↗
ORCID search ↗
Holistic Autonomous Driving Understanding by Bird'View Injected Multi-Modal Large Models
CVPR 2024
·
Xinpeng Ding
DBLP profile ↗
ORCID search ↗
HumanRefiner: Benchmarking Abnormal Human Generation and Refining with Coarse-to-Fine Pose-Reversible Guidance
ECCV 2024
·
Guian Fang
DBLP profile ↗
ORCID search ↗
Implicit Concept Removal of Diffusion Models
ECCV 2024
·
Zhili Liu
DBLP profile ↗
ORCID search ↗
Ins-DetCLIP: Aligning Detection Model to Follow Human-Language Instruction
ICLR 2024
·
Renjie Pi
DBLP profile ↗
ORCID search ↗
LayerDiff: Exploring Text-Guided Multi-layered Composable Image Synthesis via Layer-Collaborative Diffusion Model
ECCV 2024
·
Runhui Huang
DBLP profile ↗
ORCID search ↗
PanGu-Draw: Advancing Resource-Efficient Text-to-Image Synthesis with Time-Decoupled Training and Reusable Coop-Diffusion
ECCV 2024
·
Guansong Lu
DBLP profile ↗
ORCID search ↗
Reason2Drive: Towards Interpretable and Chain-Based Reasoning for Autonomous Driving
ECCV 2024
·
Ming Nie
DBLP profile ↗
ORCID search ↗
SlowFocus: Enhancing Fine-grained Temporal Understanding in Video LLM
NeurIPS 2024
·
Ming Nie
DBLP profile ↗
ORCID search ↗
UNIT: Unifying Image and Text Recognition in One Vision Encoder
NeurIPS 2024
·
Yi Zhu
DBLP profile ↗
ORCID search ↗
VidMan: Exploiting Implicit Dynamics from Video Diffusion Model for Effective Robot Manipulation
NeurIPS 2024
·
Youpeng Wen
DBLP profile ↗
ORCID search ↗