Playing Against a Stationary Opponent
Published Online:13 Jul 2026https://doi.org/10.1287/moor.2025.0904
References
- [1] (1965) Discounted dynamic programming. Ann. Math. Statist. 36(1):226–235.Crossref, Google Scholar
- [2] (1968) The big match. Ann. Math. Statist. 39(1):159–163.Crossref, Google Scholar
- [3] (2012) A continuous time approach for the asymptotic value in two-person zero-sum repeated games. SIAM J. Control Optim. 50(3):1573–1596.Crossref, Google Scholar
- [4] (2024) Solving long-run average reward robust MDPs via stochastic games. Proc. 33rd Internat. Joint Conf. Artificial Intelligence (International Joint Conferences on Artificial Intelligence, Jeju, Republic of Korea), 6707–6715.Google Scholar
- [5] (1992) Repeated games with absorbing states and no signals. Internat. J. Game Theory 21:161–174.Crossref, Google Scholar
- [6] (1997) Cyclic Markov equilibria in stochastic games. Internat. J. Game Theory 26:303–314.Crossref, Google Scholar
- [7] (1957) Stochastic games with zero stop probabilities. Dresher M, Tucker AW, Wolfe P, eds. Contributions to the Theory of Games III, Annals of Mathematics Studies, No. 39 (Princeton University Press, Princeton, NJ), 179–187.Google Scholar
- [8] (2023) Robust Markov decision processes: Beyond rectangularity. Math. Oper. Res. 48(1):203–226.Link, Google Scholar
- [9] (2026) Beyond discounted returns: Robust Markov decision processes with average and Blackwell optimality. Oper. Res., ePub ahead of print March 4, https://doi.org/10.1287/opre.2023.0694.Link, Google Scholar
- [10] (2005) Robust dynamic programming. Math. Oper. Res. 30(2):257–280.Link, Google Scholar
- [11] (1988) Finite rationality and interpersonal complexity in repeated games. Econometrica 56(2):397–410.Crossref, Google Scholar
- [12] (1974) Repeated games with absorbing states. Ann. Statist. 2(4):724–738.Google Scholar
- [13] (2010) Explicit formulas for repeated games with absorbing states. Internat. J. Game Theory 39:53–69.Crossref, Google Scholar
- [14] (1960) An approximation theorem for the Poisson binomial distribution. Pacific J. Math. 10(4):1181–1197.Crossref, Google Scholar
- [15] Le Tallec Y (2007) Robust, risk-sensitive, and data-driven control of Markov decision processes. Unpublished PhD thesis, Massachusetts Institute of Technology, Cambridge, MA.Google Scholar
- [16] (1969) Stochastic games with perfect information and time average payoff. SIAM Rev. 11(4):604–607.Crossref, Google Scholar
- [17] (1982) Stochastic games have a value. Proc. Natl. Acad. Sci. USA 79(6):2145–2146.Crossref, Google Scholar
- [18] (2009) Absorbing games with compact action spaces. Math. Oper. Res. 34(2):257–262.Link, Google Scholar
- [19] (2015) Repeated Games, Econometric Society Monographs (Cambridge University Press, New York).Crossref, Google Scholar
- [20] (1998) Finitely repeated games with finite automata. Math. Oper. Res. 23(3):513–552.Link, Google Scholar
- [21] (2005) Robust control of Markov decision processes with uncertain transition matrices. Oper. Res. 53(5):780–798.Link, Google Scholar
- [22] (1953) Stochastic games. Proc. Natl. Acad. Sci. USA 39(10):1095–1100.Crossref, Google Scholar
- [23] (1999) Three-player absorbing games. Math. Oper. Res. 24(3):669–698.Link, Google Scholar
- [24] (2007) Bounded parameter Markov decision processes with average reward criterion. Internat. Conf. Comput. Learn. Theory (Springer, Berlin, Heidelberg), 263–277.Google Scholar
- [25] (1989) On equilibria in repeated games with absorbing states. Internat. J. Game Theory 18(3):293–310.Crossref, Google Scholar
- [26] (1949) Statistical decision functions. Ann. Math. Statist. 20(2):165–205.Crossref, Google Scholar
- [27] (2023) Robust average-reward Markov decision processes. Proc. AAAI Conf. Artificial Intelligence, vol. 37, 15215–15223.Google Scholar
- [28] (2013) Robust Markov decision processes. Math. Oper. Res. 38(1):153–183.Link, Google Scholar

