Unregularized Linear Convergence in Zero-Sum Game from Preference Feedback

RL Theory

Shulun Chen, Runlong Zhou, Zihan Zhang, Maryam Fazel, Simon S. Du

We proved linear convergence of OMWU in two-player zero-sum games without assuming unique Nash equilibrium and with improved instance-dependent bounds compared to prior work.

Abstract