Results 41 to 50 of about 38,584 (203)
Bounded parameter Markov decision processes with average reward criterion [PDF]
. Bounded parameter Markov Decision Processes (BMDPs) address the issue of dealing with uncertainty in the parameters of a Markov Decision Process (MDP).
Tewari, Ambuj +3 more
core +1 more source
The risk probability optimal problem for infinite discounted semi-Markov decision processes [PDF]
summary:This paper investigates the risk probability minimization problem for infinite horizon semi-Markov decision processes (SMDPs) with varying discount factors.
Wen, Xian, Cui, Jinhua, Huo, Haifeng
core +1 more source
Software Vulnerability Patch Management with Semi-Markov Decision Process [PDF]
Information security incidents frequency has been increasing dramatically, the aim of this study is to analyze the state-space reachability problems through the transition of vulnerable status after the informative system vulnerability exposure.
Farn, Kwo-Jean +3 more
core +1 more source
An Optimal Opportunistic Maintenance Planning Integrating Discrete- and Continuous-State Information
Information-driven group maintenance is crucial to enhance the operational availability and profitability of diverse industrial systems. Existing group maintenance models have primarily concentrated on a single health criterion upon maintenance ...
Fanping Wei +4 more
doaj +1 more source
Inference Strategies for Solving Semi-Markov Decision Processes
Semi-Markov decision processes are used to formulate many control problems and also play a key role in hierarchical reinforcement learning. In this chapter we show how to translate the decision making problem into a form that can instead be solved by inference and learning techniques.
Hoffman, M, de Freitas, N
openaire +2 more sources
Decision problem for infinite duration semi-Markov process [PDF]
In the paper there are presented basic concepts and some results of the theory of semi-Markov decision processes. The optimization problem for the infinite duration SM process is connsider in the paper.
Grabski, F.
core
First passage risk probability optimality for continuous time Markov decision processes [PDF]
summary:In this paper, we study continuous time Markov decision processes (CTMDPs) with a denumerable state space, a Borel action space, unbounded transition rates and nonnegative reward function.
Wen, Xian, Huo, Haifeng
core +1 more source
Semi-Markov decision processes with partial observation [PDF]
We study a semi-Markov decision process on a Borel state space where the state of the system is not known to the decision maker but it takes decisions based on an observation process. We transform this into an equivalent problem with complete information.
Ghosh, Mrinal K, Goswami, Anindya
core
Multi-Robot Confrontation on physics-based simulators is a complex and time-consuming task, but simulators are required to evaluate the performance of the advanced algorithms.
Chunyang Hu, Meng Xu
doaj +1 more source
Multi‐agent reinforcement learning based transmission scheme for IRS‐assisted multi‐UAV systems
In this paper, a transmission scheme based on multi‐agent reinforcement learning for intelligent reflecting surface (IRS)‐assisted multiple unmanned aerial vehicles (UAVs) systems is proposed.
Yumo Mei +4 more
doaj +1 more source

