Results 21 to 30 of about 71,442 (256)
Oil tanker disasters have been a cause of major environmental disasters, with multi-generational impacts. One of the greatest hazards is damage to the propulsion system that causes the ship to turn sideways to a wave and lose stability, which in storm ...
Zbigniew Łosiewicz +4 more
doaj +1 more source
Application of the Pareto front for risk control in the transport system [PDF]
The article describes the developed model of controlling the process of means of transport operation, in which the choice of control strategy is carried out using non-deterministic methods.
Sołtysiak Agnieszka, Migawa Klaudiusz
doaj +1 more source
Towards Analysis of Semi-Markov Decision Processes [PDF]
We investigate Semi-Markov Decision Processes (SMDPs). Two problems are studied, namely, the time-bounded reachability problem and the long-run average fraction of time problem. The former aims to compute the maximal (or minimum) probability to reach a certain set of states within a given time bound.
Taolue Chen 0001, Jian Lu 0001
openaire +1 more source
A survey on semi-Markov decision processes [PDF]
This paper is a survey on semi-Markov decision processes (SMDPs). We present the background, the signi cance, and the research actuality of the in nite horizon expected discounted reward criterion, the long-run expected average reward criterion, the nite horizon expected reward criterion, the expected rst passage reward criterion, the probability ...
YongHui HUANG, XianPing GUO
openaire +1 more source
Fast Two-Stage Computation of an Index Policy for Multi-Armed Bandits with Setup Delays
We consider the multi-armed bandit problem with penalties for switching that include setup delays and costs, extending the former results of the author for the special case with no switching delays.
José Niño-Mora
doaj +1 more source
A Fast-Pivoting Algorithm for Whittle’s Restless Bandit Index
The Whittle index for restless bandits (two-action semi-Markov decision processes) provides an intuitively appealing optimal policy for controlling a single generic project that can be active (engaged) or passive (rested) at each decision epoch, and ...
José Niño-Mora
doaj +1 more source
Reactive Reinforcement Learning in Asynchronous Environments
The relationship between a reinforcement learning (RL) agent and an asynchronous environment is often ignored. Frequently used models of the interaction between an agent and its environment, such as Markov Decision Processes (MDP) or Semi-Markov Decision
Jaden B. Travnik +6 more
doaj +1 more source
Hierarchical dialogue optimization using semi-Markov decision processes [PDF]
This paper addresses the problem of dialogue optimization on large search spaces. For such a purpose, in this paper we propose to learn dialogue strategies using multiple Semi-Markov Decision Processes and hierarchical reinforcement learning. This approach factorizes state variables and actions in order to learn a hierarchy of policies. Our experiments
Cuayáhuitl, Heriberto +3 more
openaire +2 more sources
In the article a common semi-Markov mathematical model is considered that allows one to investigate the productivity and reliability of various technological processes of mechanical assembly production.
Rapatskiy Yuri +5 more
doaj +1 more source
Optimal Intervention in Semi-Markov-Based Asynchronous Probabilistic Boolean Networks
Synchronous probabilistic Boolean networks (PBNs) and generalized asynchronous PBNs have received significant attention over the past decade as a tool for modeling complex genetic regulatory networks.
Qiuli Liu +3 more
doaj +1 more source

