P
PaperPicks
Conferences
Zongqing Lu
Peking University, Beijing, China
48 papers at tracked venues · 35 at CORE A* · active 2024–2026
DBLP profile ↗
ORCID 0000-0003-3967-2704 ↗
Homepage ↗
Venues
ICLR
×11
NeurIPS
×10
ICML
×5
ECCV
×4
ICCV
×4
AAAI
×3
NAACL
×3
AAMAS
×2
ACL
×2
UAI
×2
EMNLP
×1
IROS
×1
Frequent coauthors
Jiafei Lyu
DBLP profile ↗
ORCID search ↗
×5
Haobin Jiang
DBLP profile ↗
ORCID search ↗
×4
Wanpeng Zhang
DBLP profile ↗
ORCID search ↗
×4
Bohan Zhou
DBLP profile ↗
ORCID search ↗
×3
Haoqi Yuan
DBLP profile ↗
ORCID search ↗
×3
Xiaopeng Yu
DBLP profile ↗
ORCID search ↗
×3
Hao Luo
DBLP profile ↗
ORCID search ↗
×3
Yicheng Feng
DBLP profile ↗
ORCID search ↗
×3
Jiangxing Wang
DBLP profile ↗
ORCID search ↗
×2
Kefan Su
DBLP profile ↗
ORCID search ↗
×2
Sipeng Zheng
DBLP profile ↗
ORCID search ↗
×2
Jiechuan Jiang
DBLP profile ↗
ORCID search ↗
×1
Papers
Learning Diverse Bimanual Dexterous Manipulation Skills from Human Demonstrations
AAAI 2026
·
Bohan Zhou
DBLP profile ↗
ORCID search ↗
Best Possible Q-Learning
UAI 2025
·
Jiechuan Jiang
DBLP profile ↗
ORCID search ↗
Cradle: Empowering Foundation Agents towards General Computer Control
ICML 2025
·
Weihao Tan
DBLP profile ↗
ORCID search ↗
Creative Agents: Empowering Agents with Imagination for Creative Tasks
UAI 2025
·
Penglin Cai
DBLP profile ↗
ORCID search ↗
Cross-Domain Offline Policy Adaptation with Optimal Transport and Dataset Constraint
ICLR 2025
·
Jiafei Lyu
DBLP profile ↗
ORCID search ↗
Cross-Embodiment Dexterous Grasping with Reinforcement Learning
ICLR 2025
·
Haoqi Yuan
DBLP profile ↗
ORCID search ↗
Discrete Latent Plans via Semantic Skill Abstractions
ICLR 2025
·
Haobin Jiang
DBLP profile ↗
ORCID search ↗
Efficient Residual Learning with Mixture-of-Experts for Universal Dexterous Grasping
ICLR 2025
·
Ziye Huang
DBLP profile ↗
ORCID search ↗
From Experts to a Generalist: Toward General Whole-Body Control for Humanoid Robots
NeurIPS 2025
·
Yuxuan Wang
DBLP profile ↗
ORCID search ↗
From Pixels to Tokens: Byte-Pair Encoding on Quantized Visual Modalities
ICLR 2025
·
Wanpeng Zhang
DBLP profile ↗
ORCID search ↗
GAMEBoT: Transparent Assessment of LLM Reasoning in Games
ACL 2025
·
Wenye Lin
DBLP profile ↗
ORCID search ↗
GTR: Guided Thought Reinforcement Prevents Thought Collapse in RL-Based VLM Agent Training
ICCV 2025
·
Tong Wei
DBLP profile ↗
ORCID search ↗
LLM-Based Explicit Models of Opponents for Multi-Agent Games
NAACL 2025
·
Xiaopeng Yu
DBLP profile ↗
ORCID search ↗
Learning Video-Conditioned Policy on Unlabelled Data with Joint Embedding Predictive Transformer
ICLR 2025
·
Hao Luo
DBLP profile ↗
ORCID search ↗
MEgoHand: Multimodal Egocentric Hand-Object Interaction Motion Generation
NeurIPS 2025
·
Bohan Zhou
DBLP profile ↗
ORCID search ↗
MLLM as Retriever: Interactively Learning Multimodal Retrieval for Embodied Agents
ICLR 2025
·
Junpeng Yue
DBLP profile ↗
ORCID search ↗
MotionCtrl: A Real-Time Controllable Vision-Language-Motion Model
ICCV 2025
·
Bin Cao
DBLP profile ↗
ORCID search ↗
NOLO: Navigate Only Look Once
IROS 2025
·
Bohan Zhou
DBLP profile ↗
ORCID search ↗
OpenMMEgo: Enhancing Egocentric Understanding for LMMs with Open Weights and Data
NeurIPS 2025
·
Hao Luo
DBLP profile ↗
ORCID search ↗
Planning with Quantized Opponent Models
NeurIPS 2025
·
Xiaopeng Yu
DBLP profile ↗
ORCID search ↗
Revisiting Cooperative Off-Policy Multi-Agent Reinforcement Learning
ICML 2025
·
Yueheng Li
DBLP profile ↗
ORCID search ↗
Scaling Large Motion Models with Million-Level Human Motions
ICML 2025
·
Ye Wang
DBLP profile ↗
ORCID search ↗
Taking Notes Brings Focus? Towards Multi-Turn Multimodal Dialogue Learning
EMNLP 2025
·
Jiazheng Liu
DBLP profile ↗
ORCID search ↗
Unified Multimodal Understanding via Byte-Pair Visual Encoding
ICCV 2025
·
Wanpeng Zhang
DBLP profile ↗
ORCID search ↗
VideoOrion: Tokenizing Object Dynamics in Videos
ICCV 2025
·
Yicheng Feng
DBLP profile ↗
ORCID search ↗
Watch Less, Do More: Implicit Skill Discovery for Video-Conditioned Policy
ICLR 2025
·
Jiangxing Wang
DBLP profile ↗
ORCID search ↗
AdaRefiner: Refining Decisions of Language Models with Adaptive Feedback
NAACL 2024
·
Wanpeng Zhang
DBLP profile ↗
ORCID search ↗
AuctionNet: A Novel Benchmark for Decision-Making in Large-Scale Games
NeurIPS 2024
·
Kefan Su
DBLP profile ↗
ORCID search ↗
Cross-Domain Policy Adaptation by Capturing Representation Mismatch
ICML 2024
·
Jiafei Lyu
DBLP profile ↗
ORCID search ↗
LLaMA-Rider: Spurring Large Language Models to Explore the Open World
NAACL 2024
·
Yicheng Feng
DBLP profile ↗
ORCID search ↗
Language Model Adaption for Reinforcement Learning with Natural Language Action Space
ACL 2024
·
Jiangxing Wang
DBLP profile ↗
ORCID search ↗
Learning Multi-Object Positional Relationships via Emergent Communication
AAAI 2024
·
Yicheng Feng
DBLP profile ↗
ORCID search ↗
Multi-Agent Alternate Q-Learning
AAMAS 2024
·
Kefan Su
DBLP profile ↗
ORCID search ↗
Multi-Agent Coordination via Multi-Level Communication
NeurIPS 2024
·
Gang Ding
DBLP profile ↗
ORCID search ↗
ODRL: A Benchmark for Off-Dynamics Reinforcement Learning
NeurIPS 2024
·
Jiafei Lyu
DBLP profile ↗
ORCID search ↗
Opponent Modeling based on Subgoal Inference
NeurIPS 2024
·
Xiaopeng Yu
DBLP profile ↗
ORCID search ↗
Pre-Trained Multi-Goal Transformers with Prompt Optimization for Efficient Online Adaptation
NeurIPS 2024
·
Haoqi Yuan
DBLP profile ↗
ORCID search ↗
Pre-Training Goal-based Models for Sample-Efficient Reinforcement Learning
ICLR 2024
·
Haoqi Yuan
DBLP profile ↗
ORCID search ↗
Pre-trained Visual Dynamics Representations for Efficient Policy Learning
ECCV 2024
·
Hao Luo
DBLP profile ↗
ORCID search ↗
RL-GPT: Integrating Reinforcement Learning and Code-as-policy
NeurIPS 2024
·
Shaoteng Liu
DBLP profile ↗
ORCID search ↗
Reinforcement Learning Friendly Vision-Language Model for Minecraft
ECCV 2024
·
Haobin Jiang
DBLP profile ↗
ORCID search ↗
SEABO: A Simple Search-Based Method for Offline Imitation Learning
ICLR 2024
·
Jiafei Lyu
DBLP profile ↗
ORCID search ↗
Settling Decentralized Multi-Agent Coordinated Exploration by Novelty Sharing
AAAI 2024
·
Haobin Jiang
DBLP profile ↗
ORCID search ↗
Steve-Eye: Equipping LLM-based Embodied Agents with Visual Perception in Open Worlds
ICLR 2024
·
Sipeng Zheng
DBLP profile ↗
ORCID search ↗
Tackling Non-Stationarity in Reinforcement Learning via Causal-Origin Representation
ICML 2024
·
Wanpeng Zhang
DBLP profile ↗
ORCID search ↗
Towards Understanding How to Reduce Generalization Gap in Visual Reinforcement Learning
AAMAS 2024
·
Jiafei Lyu
DBLP profile ↗
ORCID search ↗
UniCode: Learning a Unified Codebook for Multimodal Large Language Models
ECCV 2024
·
Sipeng Zheng
DBLP profile ↗
ORCID search ↗
Visual Grounding for Object-Level Generalization in Reinforcement Learning
ECCV 2024
·
Haobin Jiang
DBLP profile ↗
ORCID search ↗