Computing [(g,w)]-Optimal Policies in Discrete and Continuous Markov Programs.
ResearchPublished 1968
Shows how to compute policies that maximize a certain reasonable criterion (i.e., policies that are [(g,w)]-optimal) for the undiscounted infinite-horizon versions of two Markov decision problems--one a discrete-time model and the other a continuous-time model. The approach used is to parse the overall problem into at most three smaller problems that can be solved in sequence by the methods of linear programming or policy iteration. 33 pp. Ref
Document Details
- Copyright: RAND Corporation
- Availability: Web Only
- Year: 1968
- Pages: 33
- DOI: https://doi.org/10.7249/pubs
- Document Number: RM-5538-PR
Citation
RAND Style Manual
Chicago Manual of Style
This publication is part of the RAND research memorandum series. The research memorandum series, a product of RAND from 1948 to 1973, included working papers meant to report current results of RAND research to appropriate audiences.
This document and trademark(s) contained herein are protected by law. This representation of RAND intellectual property is provided for noncommercial use only. Unauthorized posting of this publication online is prohibited; linking directly to this product page is encouraged. Permission is required from RAND to reproduce, or reuse in another form, any of its research documents for commercial purposes. For information on reprint and reuse permissions, please visit www.rand.org/pubs/permissions.
RAND is a nonprofit institution that helps improve policy and decisionmaking through research and analysis. RAND's publications do not necessarily reflect the opinions of its research clients and sponsors.