مرکز منطقه ای اطلاع رساني علوم و فناوري - Identifying effective policies in approximate dynamic programming: Beyond regression

DocumentCode :

1912996

Title :

Identifying effective policies in approximate dynamic programming: Beyond regression

Author :

Maxwell, Matthew S. ; Henderson, Shane G. ; Topaloglu, Huseyin

Author_Institution :

Dept. of Oper. Res. & Inf. Eng., Cornell Univ., Ithaca, NY, USA

fYear :

2010

fDate :

5-8 Dec. 2010

Firstpage :

1079

Lastpage :

1087

Abstract :

Dynamic programming formulations may be used to solve for optimal policies in Markov decision processes. Due to computational complexity dynamic programs must often be solved approximately. We consider the case of a tunable approximation architecture used in lieu of computing true value functions. The standard methodology advocates tuning the approximation architecture via sample path information and regression to get a good fit to the true value function. We provide an example which shows that this approach may unnecessarily lead to poorly performing policies and suggest direct search methods to find better performing value function approximations. We illustrate this concept with an application from ambulance redeployment.

Keywords :

Markov processes; approximation theory; computational complexity; dynamic programming; emergency services; medicine; regression analysis; Markov decision process; ambulance redeployment; approximate dynamic programming formulation; computational complexity dynamic programs; optimal policies; regression; sample path information; true value function; tunable approximation architecture; Computer architecture; Dynamic programming; Function approximation; Markov processes; Medical services; Tuning;

fLanguage :

English

Publisher :

ieee

Conference_Titel :

Simulation Conference (WSC), Proceedings of the 2010 Winter

Conference_Location :

Baltimore, MD

ISSN :

0891-7736

Print_ISBN :

978-1-4244-9866-6

Type :

conf

DOI :

10.1109/WSC.2010.5679084

Filename :

5679084

Link To Document :

https://search.ricest.ac.ir/dl/search/defaultta.aspx?DTC=49&DC=1912996