PPaperPicks

Peter Stone

University of Texas at Austin, TX, USA

36 papers at tracked venues · 22 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Out-of-Distribution Generalization with a SPARC: Racing 100 Unseen Vehicles with a Single Policy
  2. The Essentials of AI for Life and Society: A Full-Scale AI Literacy Course Accessible to All
  3. Argus: A Compact and Versatile Foundation Model for Vision
  4. Deep Reinforcement Learning for Robotics: A Survey of Real-World Successes
  5. Dyn-O: Building Structured World Models with Object-Centric Representations
  6. Dyna-LfLH: Learning Agile Navigation in Dynamic Environments from Learned Hallucination
  7. Evaluating Generalization Capabilities of LLM-Based Agents in Mixed-Motive Scenarios Using Concordia
  8. FLaRe: Achieving Masterful and Adaptive Robot Policies with Large-Scale Reinforcement Learning Fine-Tuning
  9. GACL: Grounded Adaptive Curriculum Learning with Active Task and Performance Monitoring
  10. Hyperspherical Normalization for Scalable Deep Reinforcement Learning
  11. L3M+P: Lifelong Planning with Large Language Models
  12. Learning a Fast Mixing Exogenous Block MDP using a Single Trajectory
  13. Longhorn: State Space Models are Amortized Online Learners
  14. Multi-Agent Inverse Reinforcement Learning in Real World Unstructured Pedestrian Crowds
  15. PRESTO: Fast Motion Planning Using Diffusion Models Based on Key-Configuration Environment Representation
  16. Proto Successor Measure: Representing the Behavior Space of an RL Agent
  17. RLZero: Direct Policy Inference from Language Without In-Domain Supervision
  18. Reinforcement Learning Within the Classical Robotics Stack: A Case Study in Robot Soccer
  19. SimBa: Simplicity Bias for Scaling Up Parameters in Deep Reinforcement Learning
  20. The Essentials of AI for Life and Society: An AI Literacy Course for the University Community
  21. Asynchronous Task Plan Refinement for Multi-Robot Task and Motion Planning
  22. Building Minimal and Reusable Causal State Abstractions for Reinforcement Learning
  23. Dexterous Legged Locomotion in Confined 3D Spaces with Reinforcement Learning
  24. Discovering Creative Behaviors through DUPLEX: Diverse Universal Features for Policy Exploration
  25. Disentangled Unsupervised Skill Discovery for Efficient Hierarchical Reinforcement Learning
  26. LaRS: Latent Reasoning Skills for Chain-of-Thought Reasoning
  27. Learning Optimal Advantage from Preferences and Mistaking It for Reward
  28. Minimum Coverage Sets for Training Robust Ad Hoc Teamwork Agents
  29. N-agent Ad Hoc Teamwork
  30. Overview of t-DGR: A Trajectory-Based Deep Generative Replay Method for Continual Learning in Decision Making
  31. Relaxed Exploration Constrained Reinforcement Learning
  32. Rethinking Social Robot Navigation: Leveraging the Best of Two Worlds
  33. Reward (Mis)design for Autonomous Driving (Abstract Reprint)
  34. Sample Efficient Myopic Exploration Through Multitask Reinforcement Learning with Diverse Tasks
  35. SkiLD: Unsupervised Skill Discovery Guided by Factor Interactions
  36. Wait, That Feels Familiar: Learning to Extrapolate Human Preferences for Preference-Aligned Path Planning