PPaperPicks

Henghui Ding

36 papers at tracked venues · 31 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. AerialMind: Towards Referring Multi-Object Tracking in UAV Scenarios
  2. Any Detector Can Detect Anything
  3. Free-Form Scene Editor: Enabling Multi-Round Object Manipulation Like in a 3D Engine
  4. Segment Anything Across Shots: A Method and Benchmark
  5. AnyI2V: Animating Any Conditional Image with Motion Control
  6. CharaConsist: Fine-Grained Consistent Character Generation
  7. DViN: Dynamic Visual Routing Network for Weakly Supervised Referring Expression Comprehension
  8. Exploiting Temporal State Space Sharing for Video Semantic Segmentation
  9. Explore In-Context Segmentation via Latent Diffusion Models
  10. Free-Form Motion Control: Controlling the 6D Poses of Camera and Objects in Video Generation
  11. Hierarchical Alignment-enhanced Adaptive Grounding Network for Generalized Referring Expression Comprehension
  12. Hierarchical Visual Prompt Learning for Continual Video Instance Segmentation
  13. LazyDiT: Lazy Learning for the Acceleration of Diffusion Transformers
  14. MOVE: Motion-Guided Few-Shot Video Object Segmentation
  15. PVUW 2025 Challenge Report: Advances in Pixel-level Understanding of Complex Videos in the Wild
    CVPR 2025 · Henghui Ding
  16. QuartDepth: Post-Training Quantization for Real-Time Depth Estimation on the Edge
  17. ReferSplat: Referring Segmentation in 3D Gaussian Splatting
  18. SAMA: Towards Multi-Turn Referential Grounded Video Chat with Large Language Models
  19. SceneDesigner: Controllable Multi-Object Image Generation with 9-DoF Pose Manipulation
  20. Towards Omnimodal Expressions and Reasoning in Referring Audio-Visual Segmentation
  21. 3D-GRES: Generalized 3D Referring Expression Segmentation
  22. Decoupling Static and Hierarchical Motion Perception for Referring Video Segmentation
  23. Duolando: Follower GPT with Off-Policy Reinforcement Learning for Dance Accompaniment
  24. How to Continually Adapt Text-to-Image Diffusion Models for Flexible Customization?
  25. LSVOS Challenge Report: Large-Scale Complex and Long Video Object Segmentation
    ECCV 2024 · Henghui Ding
  26. Mitigating the Curse of Dimensionality for Certified Robustness via Dual Randomized Smoothing
  27. OMG-Seg: Is One Model Good Enough for all Segmentation?
  28. PVUW 2024 Challenge on Complex Video Understanding: Methods and Results
    ECCV 2024 · Henghui Ding
  29. PointCVaR: Risk-Optimized Outlier Removal for Robust 3D Point Cloud Classification
  30. RefMask3D: Language-Guided Transformer for 3D Referring Segmentation
  31. Referring Image Editing: Object-Level Image Editing via Referring Expressions
  32. Region-Native Visual Tokenization
  33. Segment Anything with Precise Interaction
  34. SemFlow: Binding Semantic Segmentation and Image Synthesis via Rectified Flow
  35. Transferable Adversarial Attacks on SAM and Its Downstream Models
  36. 🤖 SegPoint: Segment Any Point Cloud via Large Language Model