,
Richard Mayr
Creative Commons Attribution 4.0 International license
We consider the strategy complexity (i.e., memory and randomization) of optimal strategies in turn-based 2-player zero-sum stochastic games. Results in [Gimbert and Kelmendi, 2023; Richard Mayr et al., 2021] show how to lift optimal memoryless strategies for shift-invariant inverse-submixing objectives from MDPs to 2-player stochastic games with an exponential increase in the number of memory modes. We show the corresponding lower bound, i.e., the extra exponential memory is required in general, even for randomized strategies. Moreover, we solve the strategy complexity of the well-studied mean-payoff-parity objective (MP > 0 ∩ EPAR) in 2-player stochastic games. This objective is also shift-invariant inverse-submixing, but easier than the worst case for this class. In MDPs, Maximizer has optimal memoryless randomized strategies, while optimal deterministic strategies require exponential memory. However, in stochastic games, optimal randomized strategies require, at least and at most, linear memory (equal to the number of even colors). Finally, we show that the different construction in [Gimbert and Zielonka, 2009; Patricia Bouyer et al., 2023] for lifting memoryless (resp. finite-memory) deterministic strategies from MDPs (resp. 1-player games) to 2-player games cannot be generalized even to memoryless randomized strategies. We construct a shift-invariant objective where Max and Min each have optimal memoryless randomized strategies in all MDPs, but optimal (randomized) Max strategies still require infinite memory in deterministic 2-player games.
@InProceedings{dantam_et_al:LIPIcs.CONCUR.2026.30,
author = {Dantam, Mohan and Mayr, Richard},
title = {{Mean-Payoff-Parity and Lifting Strategies from MDPs to 2-Player Stochastic Games}},
booktitle = {37th International Conference on Concurrency Theory (CONCUR 2026)},
pages = {30:1--30:18},
series = {Leibniz International Proceedings in Informatics (LIPIcs)},
ISBN = {978-3-95977-447-5},
ISSN = {1868-8969},
year = {2026},
volume = {391},
editor = {Sokolova, Ana and Totzke, Patrick},
publisher = {Schloss Dagstuhl -- Leibniz-Zentrum f{\"u}r Informatik},
address = {Dagstuhl, Germany},
URL = {https://drops.dagstuhl.de/entities/document/10.4230/LIPIcs.CONCUR.2026.30},
URN = {urn:nbn:de:0030-drops-273601},
doi = {10.4230/LIPIcs.CONCUR.2026.30},
annote = {Keywords: MDPs, Stochastic Games, Parity, Mean-payoff, Strategy complexity}
}