Seminar

Past

Separation of non-ergodic uniform convergence rates for regularized learning in games

Julien Grand-Clément

  • Date9 April 2026
  • Time 11h00 - 12h15
  • Room Auditorium 5

Abstract

Self-play via online learning is a leading paradigm for solving large-scale games and has enabled recent superhuman performance (e.g., Go, Poker). This work clarifies that different convergence notions in self-play (last iterate, best iterate, and a randomly sampled iterate) can behave fundamentally differently. For a broad class of learning dynamics, including Optimistic Multiplicative Weights Update (OMWU), we prove a separation: even in two-player zero-sum games, last-iterate convergence can be arbitrarily slow, random-iterate convergence can be slower than any polynomial, while best-iterate convergence is polynomial. This departs from much prior theory where these notions align, and we attribute the gap to OMWU’s insufficient “forgetfulness,” linking it to empirical behavior in practical game solving.

Related document(s)

Other seminars

To be announced

  • Seminar

  • MAD-Stat. Seminar

  • Date 4 March 2027

  • Place Auditorium JJ Laffont

  • Speaker or organiser Agnes Lagnoux (Ecole Normale Supérieure - Université Paris Sciences & Lettres)

Details

To be announced

  • Seminar

  • MAD-Stat. Seminar

  • Date 3 December 2026

  • Place Auditorium JJ Laffont

  • Speaker or organiser Eleanor Archer (Université Paris-Dauphine)

Details

To be announced

  • Seminar

  • MAD-Stat. Seminar

  • Date 26 November 2026

  • Place A définir

  • Speaker or organiser Jason D. Hartline (Northwestern University)

Details