PPaperPicks

Bin Zhao

Northwestern Polytechnical University, School of Artificial Intelligence, Optics and Electronics, iOPEN, Xi'an, Shaanxi, China

26 papers at tracked venues · 18 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. FreeGaussian: Annotation-free Control of Articulated Objects via 3D Gaussian Splats with Flow Derivatives
  2. AerialVG: A Challenging Benchmark for Aerial Visual Grounding by Exploring Positional Relations
  3. AlignBot: Aligning VLM-Powered Customized Task Planning with User Reminders Through Fine-Tuning for Household Robots
  4. COHERENT: Collaboration of Heterogeneous Multi-Robot System with Large Language Models
  5. Cocube: a Tabletop Modular Multi-Robot Platform for Education and Research
  6. Efficient Diffusion as Low Light Enhancer
  7. Learning 2D Invariant Affordance Knowledge for 3D Affordance Grounding
  8. MoMa-Kitchen: A 100K+ Benchmark for Affordance-Grounded Last-Mile Navigation in Mobile Manipulation
  9. Open-Vocabulary Octree-Graph for 3D Scene Understanding
  10. Think Small, Act Big: Primitive Prompt Learning for Lifelong Robot Manipulation
  11. Any2Point: Empowering Any-Modality Large Models for Efficient 3D Understanding
  12. Color Event Enhanced Single-Exposure HDR Imaging
  13. Cyclic Learning for Binaural Audio Generation and Localization
  14. Depth Helps: Improving Pre-trained RGB-based Policy with Depth Information Injection
  15. GS-SLAM: Dense Visual SLAM with 3D Gaussian Splatting
  16. HPL-ESS: Hybrid Pseudo-Labeling for Unsupervised Event-based Semantic Segmentation
  17. Implicit Event-RGBD Neural SLAM
  18. Kinematic-aware Prompting for Generalizable Articulated Object Manipulation with LLMs
  19. Learning Manipulation by Predicting Interaction
  20. Learning an Actionable Discrete Diffusion Policy via Large-Scale Actionless Video Pre-Training
  21. LiveScene: Language Embedding Interactive Radiance Fields for Physical Scene Control and Rendering
  22. Point-PEFT: Parameter-Efficient Fine-Tuning for 3D Pre-trained Models
  23. Robust Quadrupedal Locomotion via Risk-Averse Policy Learning
  24. SAM-E: Leveraging Visual Foundation Model with Sequence Imitation for Embodied Manipulation
  25. TAS: Personalized Text-guided Audio Spatialization
  26. X4D-SceneFormer: Enhanced Scene Understanding on 4D Point Cloud Videos through Cross-Modal Knowledge Transfer