PPaperPicks

Matthieu Geist

16 papers at tracked venues · 12 at CORE A* · active 20242025

Venues

Frequent coauthors

Papers

  1. Learning Equilibria from Data: Provably Efficient Multi-Agent Imitation Learning
  2. Self-Improving Robust Preference Optimization
  3. ShiQ: Bringing back Bellman to LLMs
  4. Closing the Gap between TD Learning and Supervised Learning - A Generalisation Point of View
  5. Contrastive Policy Gradient: Aligning LLMs on sequence-level scores in a supervised-friendly fashion
  6. DRIFT: Deep Reinforcement Learning for Intelligent Floating Platforms Trajectories
  7. Imitating Language via Scalable Inverse Reinforcement Learning
  8. Learning Discrete-Time Major-Minor Mean Field Games
  9. MusicRL: Aligning Music Generation to Human Preferences
  10. Nash Learning from Human Feedback
  11. Near-Optimal Distributionally Robust Reinforcement Learning with General $L_p$ Norms
  12. On-Policy Distillation of Language Models: Learning from Self-Generated Mistakes
  13. Periodic agent-state based Q-learning for POMDPs
  14. Population-aware Online Mirror Descent for Mean-Field Games by Deep Reinforcement Learning
  15. Time-Constrained Robust MDPs
  16. Towards Minimax Optimality of Model-based Robust Reinforcement Learning