Exploiting the Structural Properties of the Underlying Markov Decision Problem in the Q-Learning Algorithm

Published Online:https://doi.org/10.1287/ijoc.1070.0240

Supplemental Material

ijoc.1070.0240-sm-kunnumkal_and_topaloglu.pdf (105 KB)

INFORMS site uses cookies to store information on your computer. Some are essential to make our site work; Others help us improve the user experience. By using this site, you consent to the placement of these cookies. Please read our Privacy Statement to learn more.