Back to papers
April 16, 2026cs.LG

Optimal last-iterate convergence in matrix games with bandit feedback using the log-barrier

Categories

cs.LG