Playing Against a Stationary Opponent

Published Online:https://doi.org/10.1287/moor.2025.0904

References

  • [1] Blackwell D (1965) Discounted dynamic programming. Ann. Math. Statist. 36(1):226–235.CrossrefGoogle Scholar
  • [2] Blackwell D, Ferguson TS (1968) The big match. Ann. Math. Statist. 39(1):159–163.CrossrefGoogle Scholar
  • [3] Cardaliaguet P, Laraki R, Sorin S (2012) A continuous time approach for the asymptotic value in two-person zero-sum repeated games. SIAM J. Control Optim. 50(3):1573–1596.CrossrefGoogle Scholar
  • [4] Chatterjee K, Goharshady EK, Karrabi M, Novotnỳ P, Žikelić D (2024) Solving long-run average reward robust MDPs via stochastic games. Proc. 33rd Internat. Joint Conf. Artificial Intelligence (International Joint Conferences on Artificial Intelligence, Jeju, Republic of Korea), 6707–6715.Google Scholar
  • [5] Coulomb JM (1992) Repeated games with absorbing states and no signals. Internat. J. Game Theory 21:161–174.CrossrefGoogle Scholar
  • [6] Flesch J, Thuijsman F, Vrieze K (1997) Cyclic Markov equilibria in stochastic games. Internat. J. Game Theory 26:303–314.CrossrefGoogle Scholar
  • [7] Gillette D (1957) Stochastic games with zero stop probabilities. Dresher M, Tucker AW, Wolfe P, eds. Contributions to the Theory of Games III, Annals of Mathematics Studies, No. 39 (Princeton University Press, Princeton, NJ), 179–187.Google Scholar
  • [8] Goyal V, Grand-Clément J (2023) Robust Markov decision processes: Beyond rectangularity. Math. Oper. Res. 48(1):203–226.LinkGoogle Scholar
  • [9] Grand-Clément J, Petrik M, Vieille N (2026) Beyond discounted returns: Robust Markov decision processes with average and Blackwell optimality. Oper. Res., ePub ahead of print March 4, https://doi.org/10.1287/opre.2023.0694.LinkGoogle Scholar
  • [10] Iyengar G (2005) Robust dynamic programming. Math. Oper. Res. 30(2):257–280.LinkGoogle Scholar
  • [11] Kalai E, Stanford W (1988) Finite rationality and interpersonal complexity in repeated games. Econometrica 56(2):397–410.CrossrefGoogle Scholar
  • [12] Kohlberg E (1974) Repeated games with absorbing states. Ann. Statist. 2(4):724–738.Google Scholar
  • [13] Laraki R (2010) Explicit formulas for repeated games with absorbing states. Internat. J. Game Theory 39:53–69.CrossrefGoogle Scholar
  • [14] Le Cam L (1960) An approximation theorem for the Poisson binomial distribution. Pacific J. Math. 10(4):1181–1197.CrossrefGoogle Scholar
  • [15] Le Tallec Y (2007) Robust, risk-sensitive, and data-driven control of Markov decision processes. Unpublished PhD thesis, Massachusetts Institute of Technology, Cambridge, MA.Google Scholar
  • [16] Liggett TM, Lippman SA (1969) Stochastic games with perfect information and time average payoff. SIAM Rev. 11(4):604–607.CrossrefGoogle Scholar
  • [17] Mertens JF, Neyman A (1982) Stochastic games have a value. Proc. Natl. Acad. Sci. USA 79(6):2145–2146.CrossrefGoogle Scholar
  • [18] Mertens JF, Neyman A, Rosenberg D (2009) Absorbing games with compact action spaces. Math. Oper. Res. 34(2):257–262.LinkGoogle Scholar
  • [19] Mertens JF, Sorin S, Zamir S (2015) Repeated Games, Econometric Society Monographs (Cambridge University Press, New York).CrossrefGoogle Scholar
  • [20] Neyman A (1998) Finitely repeated games with finite automata. Math. Oper. Res. 23(3):513–552.LinkGoogle Scholar
  • [21] Nilim A, El Ghaoui L (2005) Robust control of Markov decision processes with uncertain transition matrices. Oper. Res. 53(5):780–798.LinkGoogle Scholar
  • [22] Shapley LS (1953) Stochastic games. Proc. Natl. Acad. Sci. USA 39(10):1095–1100.CrossrefGoogle Scholar
  • [23] Solan E (1999) Three-player absorbing games. Math. Oper. Res. 24(3):669–698.LinkGoogle Scholar
  • [24] Tewari A, Bartlett PL (2007) Bounded parameter Markov decision processes with average reward criterion. Internat. Conf. Comput. Learn. Theory (Springer, Berlin, Heidelberg), 263–277.Google Scholar
  • [25] Vrieze OJ, Thuijsman F (1989) On equilibria in repeated games with absorbing states. Internat. J. Game Theory 18(3):293–310.CrossrefGoogle Scholar
  • [26] Wald A (1949) Statistical decision functions. Ann. Math. Statist. 20(2):165–205.CrossrefGoogle Scholar
  • [27] Wang Y, Velasquez A, Atia G, Prater-Bennette A, Zou S (2023) Robust average-reward Markov decision processes. Proc. AAAI Conf. Artificial Intelligence, vol. 37, 15215–15223.Google Scholar
  • [28] Wiesemann W, Kuhn D, Rustem B (2013) Robust Markov decision processes. Math. Oper. Res. 38(1):153–183.LinkGoogle Scholar
INFORMS site uses cookies to store information on your computer. Some are essential to make our site work; Others help us improve the user experience. By using this site, you consent to the placement of these cookies. Please read our Privacy Statement to learn more.