Results 181 to 190 of about 38,584 (203)
Some of the next articles are maybe not open access.
Deterministic policy gradient algorithms for semi‐Markov decision processes
International Journal of Intelligent Systems, 2021Ashkan Haji Hosseinloo +1 more
openaire +1 more source
Learning Automaton for Finite Semi-Markov Decision Processes
1983A finite semi-Markov decision process is studied to maximize the expected average reward. The semi-Markov kernel of the process depends on an unknown parameter taking values in a subset [a, b] of ℝS. A controller modelled as a learning automaton updates sequentially the probabilities of generating decisions based on the observed decisions, states, and ...
openaire +1 more source
Reinforcement learning with options in semi Markov decision processes
2021Treball fi de màster de: Master in Intelligent Interactive ...
openaire +1 more source
Semi-Markov Decision-Making Processes with Vector Gains
Theory of Probability & Its Applications, 1984Vinogradskaya, T. M. +2 more
openaire +3 more sources
Cost Rate Heuristics for Semi-Markov Decision Processes
1991National Research ...
Glazebrook, K.D. +2 more
openaire +1 more source
Solving Semi-Markov Decision Problems Using Average Reward Reinforcement Learning
Management Science, 1999Abhijit Gosavi +2 more
exaly

