Mean-Field Games of Speedy Information Access with Observation Costs
References
- [1] (2013) Optimal inattention to the stock market with information costs and transactions costs. Econometrica 81(4):1455–1481.Crossref, Google Scholar
- [2] (2008) Information state for Markov decision processes with network delays. 2008 47th IEEE Conf. Decision Control (IEEE, Piscataway, NJ), 3840–3847.Google Scholar
- [3] (2011) Networked Markov decision processes with delays. IEEE Trans. Automatic Control 57(4):1013–1018.Crossref, Google Scholar
- [4] (1992) Closed-loop control with delayed information. ACM SIGMETRICS Performance Evaluation Rev. 20(1):193–204.Crossref, Google Scholar
- [5] (2023) Q-learning in regularized mean-field games. Dynam. Games Appl. 13(1):89–117.Google Scholar
- [6] (1999) Markov decision processes with noise-corrupted and delayed state observations. J. Oper. Res. Soc. 50(6):660–668.Crossref, Google Scholar
- [7] (2021) Active measure reinforcement learning for observation cost minimization. Antonie L, Zadeh PM, eds. Proc. 34th Canadian Conf. Artificial Intelligence (Canadian Artificial Intelligence Association, Vancouver, British Columbia).Google Scholar
- [8] (2022) Balancing information with observation costs in deep reinforcement learning. Kiringa I, Gambs S, Kalala KH, eds. Proc. 35th Canadian Conf. Artificial Intelligence (Canadian Artificial Intelligence Association, Toronto, Ontario).Google Scholar
- [9] (1978) An introductory approach to duality in optimal stochastic control. SIAM Rev. 20(1):62–78.Crossref, Google Scholar
- [10] (2007) Measure Theory, vol. 1 (Springer, Berlin, Heidelberg).Crossref, Google Scholar
- [11] (2009) Impulse control problem on finite horizon with execution delay. Stochastic Processes Appl. 119(5):1436–1469.Crossref, Google Scholar
- [12] (2006) Large population stochastic dynamic games: Closed-loop McKean-Vlasov systems and the Nash certainty equivalence principle. Comm. Inform. Systems 6(3):221–252.Crossref, Google Scholar
- [13] (2018) Mean field game of controls and an application to trade crowding. Math. Financial Econom. 12:335–363.Crossref, Google Scholar
- [14] (2013) Mean field forward-backward stochastic differential equations. Electron. Comm. Probab. 18:1–15.Crossref, Google Scholar
- [15] (2018) Probabilistic Theory of Mean Field Games with Applications I and II, Probability Theory and Stochastic Modelling, vol. 83 (Springer, Cham, Switzerland).Crossref, Google Scholar
- [16] (2023) Optimal execution with stochastic delay. Finance Stochastics 27(1):1–47.Crossref, Google Scholar
- [17] (2021) Delay-aware model-based reinforcement learning for continuous control. Neurocomputing 450:119–128.Crossref, Google Scholar
- [18] (2021) Modelling COVID-19 contagion: Risk assessment and targeted mitigation policies. Roy. Soc. Open Sci. 8(3):201535.Crossref, Google Scholar
- [19] (1971) An optimal stochastic control problem with observation cost. IEEE Trans. Automatic Control 16(2):185–189.Crossref, Google Scholar
- [20] (2021) Approximately solving mean field games via entropy-regularized deep reinforcement learning. Banerjee A, Fukumizu K, eds. Proc. 24th Internat. Conf. Artificial Intelligence Statist., vol. 130 (PMLR, New York), 1909–1917.Google Scholar
- [21] (2009) Large Deviations Techniques and Applications, Stochastic Modelling and Applied Probability, vol. 38 (Springer, Berlin, Heidelberg).Google Scholar
- [22] (1960) Foundations of Modern Analysis (Academic Press, New York).Google Scholar
- [23] (2019) A theory of regularized Markov decision processes. Chaudhuri K, Salakhutdinov R, eds. Proc. 36th Internat. Conf. Machine Learn., vol. 97 (PMLR, New York), 2160–2169.Google Scholar
- [24] (2011) Gibbs Measures and Phase Transitions (Walter de Gruyter, Berlin).Crossref, Google Scholar
- [25] (2016) Extended deterministic mean-field games. SIAM J. Control Optim. 54(2):1030–1055.Crossref, Google Scholar
- [26] (2026) On information controls. Preprint, submitted February 7, https://arxiv.org/abs/2602.07318.Google Scholar
- [27] (2021) Optimal causal rate-constrained sampling for a class of continuous Markov processes. IEEE Trans. Inform. Theory 67(12):7876–7890.Crossref, Google Scholar
- [28] (2024) MF-OMO: An optimization formulation of mean-field games. SIAM J. Control Optim. 62(1):243–270.Crossref, Google Scholar
- [29] (2019) Learning mean-field games. Wallach HM, Larochelle H, Beygelzimer A, d’Alché-Buc F, Fox EB, eds. Proc. 33rd Internat. Conf. Neural Inform. Processing Systems (Curran Associates Inc., Red Hook, NY), 4966–4976.Google Scholar
- [30] (2023a) A general framework for learning mean-field games. Math. Oper. Res. 48(2):656–686.Link, Google Scholar
- [31] (2023b) MFGLib: A library for mean field games. Preprint, submitted April 17, https://arxiv.org/abs/2304.08630.Google Scholar
- [32] (2008) Paging and registration in cellular networks: Jointly optimal policies and an iterative algorithm. IEEE Trans. Inform. Theory 54(2):608–622.Crossref, Google Scholar
- [33] (1989) Adaptive Markov Control Processes, Applied Mathematical Sciences, vol. 83 (Springer, New York).Crossref, Google Scholar
- [34] (1996) Discrete-Time Markov Control Processes, Applications of Mathematics, vol. 30 (Springer, New York).Crossref, Google Scholar
- [35] (2021) Self-triggered Markov decision processes. 2021 60th IEEE Conf. Decision Control (IEEE, Piscataway, NJ), 4507–4514.Google Scholar
- [36] (2003) Markov decision processes with delays and asynchronous cost collection. IEEE Trans. Automatic Control 48(4):568–574.Crossref, Google Scholar
- [37] (2025) A Fisher–Rao gradient flow for entropy-regularised Markov decision processes in Polish spaces. Foundations Comput. Math., ePub ahead of print August 11, https://doi.org/10.1007/s10208-025-09729-3. Crossref, Google Scholar
- [38] (2020) Active reinforcement learning: Observing rewards at a cost. Preprint, submitted November 13, https://arxiv.org/abs/2011.06709.Google Scholar
- [39] (2007) Mean field games. Japan J. Math. 2(1):229–260.Crossref, Google Scholar
- [40] (2022) Convergence of large population games to mean field games with interaction through the controls. SIAM J. Math. Anal. 54(3):3535–3574.Crossref, Google Scholar
- [41] (2022) Learning mean field games: A survey. Preprint, submitted May 25, https://arxiv.org/abs/2205.12944.Google Scholar
- [42] (2021) Revisiting state augmentation methods for reinforcement learning with stochastic delays. CIKM’21: Proc. 30th ACM Internat. Conf. Inform. Knowledge Management (Association for Computing Machinery, New York), 1346–1355.Google Scholar
- [43] (2013) Optimal strategies for communication and remote estimation with an energy harvesting sensor. IEEE Trans. Automatic Control 58(9):2246–2260.Crossref, Google Scholar
- [44] (2008) Optimal stochastic impulse control with delayed reaction. Appl. Math. Optim. 58(2):243–255.Crossref, Google Scholar
- [45] (2022) Scaling mean field games by online mirror descent. AAMAS’22: Proc. 21st Internat. Conf. Autonomous Agents Multiagent Systems (International Foundation for Autonomous Agents and Multiagent Systems, Richland, SC), 1028–1037.Google Scholar
- [46] (2020) Fictitious play for mean field games: Continuous time analysis and applications. Larochelle H, Ranzato M, Hadsell R, Balcan MF, Lin H, eds. NIPS’20: Proc. 34th Internat. Conf. Neural Inform. Processing Systems (Curran Associates Inc., Red Hook, NY), 13199–13213.Google Scholar
- [47] (2025) Markov decision processes with observation costs: Framework and computation with a penalty scheme. Math. Oper. Res. 50(2):1305–1332.Link, Google Scholar
- [48] (2020) Effect of delay in diagnosis on transmission of COVID-19. Math. Biosci. Engrg. 17(3):2725–2740.Crossref, Google Scholar
- [49] (2018) Markov–Nash equilibria in mean-field games with discounted cost. SIAM J. Control Optim. 56(6):4256–4287.Crossref, Google Scholar
- [50] (2019) Approximate Nash equilibria in partially observed stochastic games with mean-field interactions. Math. Oper. Res. 44(3):1006–1033.Link, Google Scholar
- [51] (2023) Partially observed discrete-time risk-sensitive mean field games. Dynam. Games Appl. 13(3):926–960.Crossref, Google Scholar
- [52] (2019) Stochastic control with delayed information and related nonlinear master equation. SIAM J. Control Optim. 57(1):693–717.Crossref, Google Scholar
- [53] (2010) Control delay in reinforcement learning for real-time dynamic systems: A memoryless approach. 2010 IEEE/RSJ Internat. Conf. Intelligent Robots Systems (IEEE, Piscataway, NJ), 3226–3231.Google Scholar
- [54] (2020) LQG control and sensing co-design. IEEE Trans. Automatic Control 66(4):1468–1483.Crossref, Google Scholar
- [55] (2014) Markov control processes with rare state observation: Theory and application to treatment scheduling in HIV–1. Comm. Math. Sci. 12(5):859–877.Crossref, Google Scholar
- [56] (2001) Asymptotic Approximations of Integrals, Classics in Applied Mathematics (Society for Industrial and Applied Mathematics, Philadelphia).Crossref, Google Scholar
- [57] (2008) Optimal sensor querying: Markovian and LQG models with controlled observations. IEEE Trans. Automatic Control 53(6):1392–1405.Crossref, Google Scholar
- [58] (2020) Analysis and computation of an optimality equation arising in an impulse control problem with discrete and costly observations. J. Comput. Appl. Math. 366:112399.Crossref, Google Scholar
- [59] (2020a) A hybrid stochastic river environmental restoration modeling with discrete and costly observations. Optimal Control Appl. Methods 41(6):1964–1994.Crossref, Google Scholar
- [60] (2021) Cost-efficient monitoring of continuous-time stochastic processes based on discrete observations. Appl. Stochastic Models Bus. Indust. 37(1):113–138.Crossref, Google Scholar
- [61] (2020b) Analysis and computation of a discrete costly observation model for growth estimation and management of biological resources. Comput. Math. Appl. 79(4):1072–1093.Crossref, Google Scholar

