Sufficiency of Markov Policies for Continuous-Time Jump Markov Decision Processes | PERPUSTAKAAN UNIVERSITAS KATOLIK PARAHYANGAN

Pencarian berdasarkan :

Pencarian terakhir:

Text

Sufficiency of Markov Policies for Continuous-Time Jump Markov Decision Processes

Feinberg, Eugene A. - Nama Orang; Mandava, Manasa - Nama Orang; Shiryaev, Albert N. - Nama Orang;

One of the basic facts known for discrete-time Markov decision processes is that, if the probability distribution of an initial state is fixed, then for every policy it is easy to construct a (randomized) Markov policy with the same marginal distributions of state-action pairs as for the original policy. This equality of marginal distributions implies that the values of major objective criteria, including expected discounted total costs and average rewards per unit time, are equal for these two policies. This paper investigates the validity of the similar fact for continuous-time jump Markov decision processes (CTJMDPs). It is shown in this paper that the equality of marginal distributions takes place for a CTJMDP if the corresponding Markov policy defines a nonexplosive jump Markov process. If this Markov process is explosive, then at each time instance, the marginal probability, that a state-action pair belongs to a measurable set of state-action pairs, is not greater for the described Markov policy than the same probability for the original policy. These results are applied in this paper to CTJMDPs with expected discounted total costs and average costs per unit time. It is shown for these criteria that, if the initial state distribution is fixed, then for every policy, there exists a Markov policy with the same or better value of the objective function.

Ketersediaan

Barcode		Tipe Koleksi	Nomor Panggil	Lokasi	Status
art142721	null	Artikel		Gdg9-Lt3	Tersedia namun tidak untuk dipinjamkan - No Loan

Informasi Detail

Judul Seri: MATHEMATICS OF OPERATIONS RESEARCH; Vol.47 No.2 May 2022
No. Panggil: -
Penerbit: : .,
Deskripsi Fisik: p. 1266-1286
Bahasa: English
ISBN/ISSN: -
Klasifikasi: NONE
Tipe Isi: -
Tipe Media: -
Tipe Pembawa: -
Edisi: -
Subjek: STATE
ACTION
CONTINUOUS-TIME JUMP MARKOV PROCESS
BOREL
MARKOV POLICY
Info Detail Spesifik: https://doi.org/10.1287/moor.2021.1169
Pernyataan Tanggungjawab: Eugene A. Feinberg, Manasa Mandava, Albert N. Shiryaev

Versi lain/terkait

Tidak tersedia versi lain

Lampiran Berkas

Tidak Ada Data

Komentar

Anda harus masuk sebelum memberikan komentar