O(T-1 Convergence of Optimistic-Follow-the-Regularized-Leader in Two-Player Zero-Sum Markov Games

Yuepeng Yang, Cong Ma
2/12/2026
Semantic Scholar

Code Implementations

No confident code match yet

We couldn't find an author-owned or strongly-evidenced community implementation for this paper. 2 weaker matches are hidden by default — verify before relying on them.

No code implementations found yet.

Know of an implementation? Let us know in the comments below!

Cite this paper

@article{yang2026ot,
  title  = {O(T-1 Convergence of Optimistic-Follow-the-Regularized-Leader in Two-Player Zero-Sum Markov Games},
  author = {Yuepeng Yang and Cong Ma},
  year   = {2026},
  url    = {https://api.semanticscholar.org/CorpusID:efc026ace73425a4f990ad5a521e8653df85d072},
  journal = {ICLR 2023 2023}
}

Discussion