Results 11 to 20 of about 9,481,834 (176)

Inference Strategies for Solving Semi-Markov Decision Processes

open access: yes, 2012
Semi-Markov decision processes are used to formulate many control problems and also play a key role in hierarchical reinforcement learning. In this chapter we show how to translate the decision making problem into a form that can instead be solved by inference and learning techniques.
Hoffman, M, de Freitas, N
core   +4 more sources

The exponential cost optimality for finite horizon semi-Markov decision processes [PDF]

open access: yes, 2022
summary:This paper considers an exponential cost optimality problem for finite horizon semi-Markov decision processes (SMDPs). The objective is to calculate an optimal policy with minimal exponential costs over the full set of policies in a finite ...
Wen, Xian, Huo, Haifeng
core   +1 more source

Optimal maintenance of deteriorating equipment using semi-Markov decision processes and linear programming [PDF]

open access: yesInternational Journal of Industrial Engineering and Management
This paper considers a mathematical model analysing the deterioration of system equipment and available maintenance options. Under specific conditions on costs and transition probabilities of the model, the issue of ideal maintenance of the equipment by ...
Giannis Kechagias   +3 more
doaj   +1 more source

Risk probability optimization problem for finite horizon continuous time Markov decision processes with loss rate [PDF]

open access: yes, 2021
summary:This paper presents a study the risk probability optimality for finite horizon continuous-time Markov decision process with loss rate and unbounded transition rates.
Wen, Xian, Huo, Haifeng
core   +1 more source

A Fast-Pivoting Algorithm for Whittle’s Restless Bandit Index

open access: yesMathematics, 2020
The Whittle index for restless bandits (two-action semi-Markov decision processes) provides an intuitively appealing optimal policy for controlling a single generic project that can be active (engaged) or passive (rested) at each decision epoch, and ...
José Niño-Mora
doaj   +1 more source

Recursive Markov Decision Processes and Recursive Stochastic Games [PDF]

open access: yes, 2005
We introduce Recursive Markov Decision Processes (RMDPs) and Recursive Simple Stochastic Games (RSSGs), and study the decidability and complexity of algorithms for their analysis and verification.
Mihalis Yannakakis   +3 more
core   +1 more source

Reactive Reinforcement Learning in Asynchronous Environments

open access: yesFrontiers in Robotics and AI, 2018
The relationship between a reinforcement learning (RL) agent and an asynchronous environment is often ignored. Frequently used models of the interaction between an agent and its environment, such as Markov Decision Processes (MDP) or Semi-Markov Decision
Jaden B. Travnik   +6 more
doaj   +1 more source

Symbolic Magnifying Lens Abstraction in Markov Decision Processes [PDF]

open access: yes, 2008
In this paper, we combine abstraction-refinement and symbolic techniques to fight the state-space explosion problem when model checking Markov decision processes (MDPs).
Luca de Alfaro   +7 more
core   +1 more source

Research of reliability and efficiency of technological processes of mechanical assembly production on the basis of the common semi-Markov model

open access: yesMATEC Web of Conferences, 2018
In the article a common semi-Markov mathematical model is considered that allows one to investigate the productivity and reliability of various technological processes of mechanical assembly production.
Rapatskiy Yuri   +5 more
doaj   +1 more source

Hierarchical dialogue optimization using semi-Markov decision processes [PDF]

open access: yesInterspeech 2007, 2007
This paper addresses the problem of dialogue optimization on large search spaces. For such a purpose, in this paper we propose to learn dialogue strategies using multiple Semi-Markov Decision Processes and hierarchical reinforcement learning. This approach factorizes state variables and actions in order to learn a hierarchy of policies. Our experiments
Cuayáhuitl, Heriberto   +3 more
openaire   +3 more sources

Home - About - Disclaimer - Privacy