Semi-infinite weighted Markov decision processes with perturbation

Abbad, Mohammed; Rahhali, Khalid

doi:10.1007/s001860400363

Semi-infinite weighted Markov decision processes with perturbation

Published: October 2004

Volume 60, pages 251–265, (2004)
Cite this article

Mathematical Methods of Operations Research Aims and scope Submit manuscript

Mohammed Abbad¹ &
Khalid Rahhali¹

46 Accesses
Explore all metrics

Abstract.

In this paper, Weighted reward Perturbed Markov Decision Processes with finite state and countable action spaces (semi-infinite WMDP for short) are considered. The ”weighted reward” refers to appropriately normalized convex combination of the discounted and the long-run average reward criteria. This criterion allows the controller to trade-off short-term rewards versus long-run rewards. In every application where both the discounted and the long-run average criteria have been proposed in the past, there is clearly a rationale for considering the weighted criterion. Of course, as with all Markov decision models, the standard weighted criterion model assumes that all the transition probabilities are known precisely. Since, in most applications this would not be the case, we consider the perturbed version of the weighted reward model. In the case of perturbations, we prove that for many models a nearly optimal strategy can be found in the class of relatively “simple ultimately deterministic” strategies. These are strategies which behave just like deterministic stationary strategies, after a certain point of time.

This is a preview of subscription content, log in via an institution to check access.

Access this article

Log in via an institution

Subscribe and save

Springer+ Basic

$34.99 /Month

Get 10 units per month
Download Article/Chapter or eBook
1 Unit = 1 Article or 1 Chapter
Cancel anytime

Buy Now

Price excludes VAT (USA)
Tax calculation will be finalised during checkout.

Instant access to the full article PDF.

Institutional subscriptions

Author information

Authors and Affiliations

Département de Mathématiques et Informatique, Faculté des Sciences, Université Mohammed V Agdal, BP, 1014, Rabat, Morocco
Mohammed Abbad & Khalid Rahhali

Authors

Mohammed Abbad
View author publications
You can also search for this author in PubMed Google Scholar
Khalid Rahhali
View author publications
You can also search for this author in PubMed Google Scholar

Corresponding author

Correspondence to Mohammed Abbad.

Additional information

Manuscript received: September 2003/Final version received: January 2004

Rights and permissions

Reprints and permissions

About this article

Cite this article

Abbad, M., Rahhali, K. Semi-infinite weighted Markov decision processes with perturbation. Math Meth Oper Res 60, 251–265 (2004). https://doi.org/10.1007/s001860400363

Download citation

Issue Date: October 2004
DOI: https://doi.org/10.1007/s001860400363

Keywords

Access this article

Log in via an institution

Subscribe and save

Springer+ Basic

$34.99 /Month

Get 10 units per month
Download Article/Chapter or eBook
1 Unit = 1 Article or 1 Chapter
Cancel anytime

Buy Now

Price excludes VAT (USA)
Tax calculation will be finalised during checkout.

Instant access to the full article PDF.

Institutional subscriptions

Semi-infinite weighted Markov decision processes with perturbation

Abstract.

Access this article

Subscribe and save

Buy Now

Similar content being viewed by others

On the Expected Total Reward with Unbounded Returns for Markov Decision Processes

The risk probability criterion for discounted continuous-time Markov decision processes

Constrained Continuous-Time Markov Decision Processes on the Finite Horizon

Author information

Authors and Affiliations

Corresponding author

Additional information

Rights and permissions

About this article

Cite this article

Keywords

Subscribe and save

Buy Now

Semi-infinite weighted Markov decision processes with perturbation

Abstract.

Access this article

Subscribe and save

Buy Now

Similar content being viewed by others

On the Expected Total Reward with Unbounded Returns for Markov Decision Processes

The risk probability criterion for discounted continuous-time Markov decision processes

Constrained Continuous-Time Markov Decision Processes on the Finite Horizon

Author information

Authors and Affiliations

Corresponding author

Additional information

Rights and permissions

About this article

Cite this article

Share this article

Keywords

Subscribe and save

Buy Now