Results 21 to 30 of about 5,928 (266)
Solving Hidden-Semi-Markov-Mode Markov Decision Problems
International audienceHidden-Mode Markov Decision Processes (HM-MDPs) were proposed to represent sequential decision-making problems in non-stationary environments that evolve according to a Markov chain.
Hadoux, Emmanuel +2 more
core +2 more sources
Markov Decision Processes with Incomplete Information and Semi-Uniform Feller Transition Probabilities [PDF]
This paper deals with control of partially observable discrete-time stochastic systems. It introduces and studies Markov Decision Processes with Incomplete Information and with semi-uniform Feller transition probabilities.
Feinberg, Eugene A. +2 more
core +1 more source
Discounted semi-Markov decision processes : linear programming and policy iteration [PDF]
For semi-Markov decision processes with discounted rewards we derive the well known results regarding the structure of optimal strategies (nonrandomized, stationary Markov strategies) and the standard algorithms (linear programming, policy iteration ...
van Nunen, J.A.E.E., Wessels, J.
core +2 more sources
Learning to maximize reward rate: a model based on semi-Markov decision processes
When animals have to make a number of decisions during a limited time interval, they face a fundamental problem: how much time they should spend on each decision in order to achieve the maximum possible total outcome.
Arash eKhodadadi +2 more
doaj +1 more source
Coordinating multiple unmanned aerial vehicle (UAV) systems under strict energy and temporal constraints remains a complex scheduling problem. Existing reinforcement learning methods typically rely on fixed-time-step modeling, which struggles to ...
Feng Wang +4 more
doaj +1 more source
This article deals with the modeling of the processes of operating both marine main and auxiliary engines. The paper presents a model of changes in operating conditions of ship’s internal combustion engine.
Landowski Bogdan +3 more
doaj +1 more source
Although Deep Reinforcement Learning (DRL) offers adaptive control for manufacturing processes, training agents within high-fidelity commercial Discrete Event Simulation (DES) tools is hindered by the temporal mismatch between the continuous-time event ...
Varici, Merve Demir +4 more
doaj +1 more source
The risk probability optimal problem for infinite discounted semi-Markov decision processes [PDF]
summary:This paper investigates the risk probability minimization problem for infinite horizon semi-Markov decision processes (SMDPs) with varying discount factors.
Wen, Xian, Cui, Jinhua, Huo, Haifeng
core +1 more source
To address the complexity and uncertainty inherent in the post-earthquake damage state recovery process of shield tunnels, a method for establishing performance recovery models based on semi-Markov processes is proposed, enabling quantitative assessment ...
ZENG Nianchen 1, HUANG Zhongkai 1, ZHANG Dongmei 1, 2, 3, GAN Binlin 1, 2
doaj +1 more source
Fostering Innovation: Streamlining Magnetocaloric Materials Research by Digitalization
Magnetocaloric cooling (MCE) is an environmentally friendly refrigeration method with great potential. Optimizing MCE materials involves the preparation and screening of large quantities of samples, which in turn generates a large amount of data. A digitalization approach is presented that uses ontologies, knowledge graphs, and digital workflows to ...
Simon Bekemeier +17 more
wiley +1 more source

