Results 181 to 190 of about 38,584 (203)
Some of the next articles are maybe not open access.

Deterministic policy gradient algorithms for semi‐Markov decision processes

International Journal of Intelligent Systems, 2021
Ashkan Haji Hosseinloo   +1 more
openaire   +1 more source

Learning Automaton for Finite Semi-Markov Decision Processes

1983
A finite semi-Markov decision process is studied to maximize the expected average reward. The semi-Markov kernel of the process depends on an unknown parameter taking values in a subset [a, b] of ℝS. A controller modelled as a learning automaton updates sequentially the probabilities of generating decisions based on the observed decisions, states, and ...
openaire   +1 more source

Reinforcement learning with options in semi Markov decision processes

2021
Treball fi de màster de: Master in Intelligent Interactive ...
openaire   +1 more source

Semi-Markov Decision-Making Processes with Vector Gains

Theory of Probability & Its Applications, 1984
Vinogradskaya, T. M.   +2 more
openaire   +3 more sources

Cost Rate Heuristics for Semi-Markov Decision Processes

1991
National Research ...
Glazebrook, K.D.   +2 more
openaire   +1 more source

Solving Semi-Markov Decision Problems Using Average Reward Reinforcement Learning

Management Science, 1999
Abhijit Gosavi   +2 more
exaly  

Home - About - Disclaimer - Privacy