PPaperPicks

Tuomas Sandholm

Carnegie Mellon University, Pittsburgh, USA

28 papers at tracked venues · 17 at CORE A* · active 20242026

Venues

Frequent coauthors

Papers

  1. Faster Game Solving via Hyperparameter Schedules
  2. Weakest Bidder Types and New Core-Selecting Combinatorial Auctions
  3. A Multiagent Path Search Algorithm for Large-Scale Coalition Structure Generation
  4. AlphaZeroES: Direct Score Maximization Outperforms Planning Loss Minimization
  5. ApproxED: Approximate Exploitability Descent via Learned Best Responses
  6. Computing Game Symmetries and Equilibria That Respect Them
  7. Expected Variational Inequalities
  8. Increasing Revenue in Efficient Combinatorial Auctions by Learning to Generate Artificial Competition
  9. Joint-Perturbation Simultaneous Pseudo-Gradient
  10. New Sequence-Independent Lifting Techniques for Cover Inequalities and When They Induce Facets
  11. The Complexity of Symmetric Equilibria in Min-Max Optimization and Team Zero-Sum Games
  12. The Value of Recall in Extensive-Form Games
  13. A Multiagent Path Search Algorithm for Large-Scale Coalition Structure Generation
  14. Automated Design of Affine Maximizer Mechanisms in Dynamic Settings
  15. Confronting Reward Model Overoptimization with Constrained RLHF
  16. Convergence of $\text{log}(1/\epsilon)$ for Gradient-Based Algorithms in Zero-Sum Games without the Condition Number: A Smoothed Analysis
  17. Efficient $\Phi$-Regret Minimization with Low-Degree Swap Deviations in Extensive-Form Games
  18. Efficient Size-based Hybrid Algorithm for Optimal Coalition Structure Generation
  19. Exponential Lower Bounds on the Double Oracle Algorithm in Zero-Sum Games
  20. Faster Optimal Coalition Structure Generation via Offline Coalition Selection and Graph-Based Search
  21. Game-Theoretic Robust Reinforcement Learning Handles Temporally-Coupled Perturbations
  22. Imperfect-Recall Games: Equilibrium Concepts and Their Complexity
  23. Mediator Interpretation and Faster Learning Algorithms for Linear Correlated Equilibria in General Sequential Games
  24. Model-Free Preference Elicitation
  25. On the Outcome Equivalence of Extensive-Form and Behavioral Correlated Equilibria
  26. Optimistic Policy Gradient in Multi-Player Markov Games with a Single Controller: Convergence beyond the Minty Property
  27. Scalable Mechanism Design for Multi-Agent Path Finding
  28. Toward Optimal Policy Population Growth in Two-Player Zero-Sum Games