PPaperPicks

Han Zhao

13 papers at tracked venues · 6 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. ReconVLA: Reconstructive Vision-Language-Action Model as Effective Robot Perceiver
  2. VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model
  3. Cobra: Extending Mamba to Multi-Modal Large Language Model for Efficient Inference
    AAAI 2025 · Han Zhao
  4. MoRE: Unlocking Scalability in Reinforcement Learning for Quadruped Vision-Language-Action Models
    ICRA 2025 · Han Zhao
  5. PD-VLA: Accelerating Vision-Language-Action Model Integrated with Action Chunking via Parallel Decoding
  6. Quart-Online: Latency-Free Multimodal Large Language Model for Quadruped Robot Learning
  7. ReinboT: Amplifying Robot Visual-Language Manipulation with Reinforcement Learning
  8. SSR: Enhancing Depth Perception in Vision-Language Models via Rationale-Guided Spatial Reasoning
  9. VLAS: Vision-Language-Action Model with Speech Instructions for Customized Robot Manipulation
  10. GeRM: A Generalist Robotic Model with Mixture-of-experts for Quadruped Robot
  11. PiTe: Pixel-Temporal Alignment for Large Video-Language Model
  12. QUAR-VLA: Vision-Language-Action Model for Quadruped Robots
  13. RL2AC: Reinforcement Learning-based Rapid Online Adaptive Control for Legged Robot Robust Locomotion