ms·1979년 4월 1일
Linear Programming and Markov Decision Chains
Arie Hordijk, L. C. M. Kallenberg
Management Science
146
피인용
4.1
FWCI
0
IS/마케팅/OM 탑저널 피인용
9
IS/마케팅/OM 탑저널 참고문헌
- 주제동적계획과 확률최적화 · 생산·최적화
01Abstract
In this paper we show that for a finite Markov decision process an average optimal policy can be found by solving only one linear programming problem. Also the relation between the set of feasible solutions of the linear program and the set of stationary policies is analyzed.
02연구 흐름
불러오는 중…
03비슷한 논문
불러오는 중…
04이후 연구
불러오는 중…
05선행 연구
불러오는 중…
06서지 정보
- 저널Management Science · 25(4) · 352–362
- 토픽Reinforcement Learning in Robotics · Artificial Intelligence
- DOI10.1287/mnsc.25.4.352
- 저자Arie Hordijk, L. C. M. Kallenberg