IS Atlas
ms·1973년 3월 1일

Semi-Markov Decision Processes with Unbounded Rewards

Steven A. Lippman

Management Science

138
피인용
17.0
FWCI
7
IS/마케팅/OM 탑저널 피인용
13
IS/마케팅/OM 탑저널 참고문헌
01Abstract

We consider a semi-Markov decision process with arbitrary action space; the state space is the nonnegative integers. As in queueing systems, we assume that {0, 1, 2, …, n + N} is the set of states accessible from state n in one transition, where N is finite and independent of n. The novel feature of this model is that the one-period reward is not required to be uniformly bounded; instead, we merely assume it to be bounded by a polynomial in n. Our main concern is with the average cost problem. A set of conditions sufficient for there to be an optimal stationary policy which can be obtained from the usual functional equation is developed. These conditions are quite weak and, as illustrated in several queueing examples, are easily verified.

02연구 흐름

불러오는 중…

03비슷한 논문

불러오는 중…

04이후 연구

불러오는 중…

05선행 연구

불러오는 중…

06서지 정보