PPaperPicks

Xiaoshuai Hao

29 papers at tracked venues · 22 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. NavA³: Understanding Any Instruction, Navigating Anywhere, Finding Anything
  2. Question-Adaptive Graph Learning for Multi-hop Retrieval Augmented Generation
  3. What You See Is What You Reach: Towards Spatial Navigation with High-Level Human Instructions
  4. AffordGrasp: In-Context Affordance Reasoning for Open-Vocabulary Task-Oriented Grasping in Clutter
  5. FastRSR: Efficient and Accurate Road Surface Reconstruction in Bird's Eye View
  6. KALAHash: Knowledge-Anchored Low-Resource Adaptation for Deep Hashing
  7. M3-Net: A Cost-Effective Graph-Free MLP-Based Model for Traffic Prediction
  8. MFFI: Multi-Dimensional Face Forgery Image Dataset for Real-World Scenarios
  9. MapNav: A Novel Memory Representation via Annotated Semantic Maps for VLM-based Vision-and-Language Navigation
  10. Open-Vocabulary Fine-Grained Hand Action Detection
  11. Reason-RFT: Reinforcement Fine-Tuning for Visual Reasoning of Vision Language Models
  12. RoboAfford: A Dataset and Benchmark for Enhancing Object and Spatial Affordance Learning in Robot Manipulation
  13. RoboBrain: A Unified Brain Model for Robotic Manipulation from Abstract to Concrete
  14. SSTAG: Structure-Aware Self-Supervised Learning Method for Text-Attributed Graphs
  15. SVD: Spatial Video Dataset
    ACM MM 2025 ·
    Mohammad Hossein Izadimehr
  16. SafeMap: Robust HD Map Construction from Incomplete Observations
    ICML 2025 · Xiaoshuai Hao
  17. Synergistic Prompting for Robust Visual Recognition with Missing Modalities
  18. TASAR: Transfer-based Attack on Skeletal Action Recognition
  19. Training-Free Generation of Temporally Consistent Rewards from VLMs
  20. VQualA 2025 Challenge on Engagement Prediction for Short Videos: Methods and Results
  21. VQualA 2025 Challenge on GenAI-Bench AIGC Video Quality Assessment: Methods and Results
  22. VQualA 2025 Challenge on Image Super-Resolution Generated Content Quality Assessment: Methods and Results
  23. Video-CoT: A Comprehensive Dataset for Spatiotemporal Understanding of Videos Based on Chain-of-Thought
  24. What Really Matters for Robust Multi-Sensor HD Map Construction?
    IROS 2025 · Xiaoshuai Hao
  25. Enhancing 3D Hand Pose Estimation via Dense Ordinal Regression Network
  26. FTF-ER: Feature-Topology Fusion-Based Experience Replay Method for Continual Graph Learning
  27. Is Your HD Map Constructor Reliable under Sensor Corruptions?
    NeurIPS 2024 · Xiaoshuai Hao
  28. MBFusion: A New Multi-modal BEV Feature Fusion Method for HD Map Construction
    ICRA 2024 · Xiaoshuai Hao
  29. MapDistill: Boosting Efficient Camera-Based HD Map Construction via Camera-LiDAR Fusion Model Distillation
    ECCV 2024 · Xiaoshuai Hao