Results 101 to 110 of about 13,577 (270)

Exploration and Exploitation: A Study on Sample Efficiency in Reinforcement Learning With Multifaceted Curiosity Rewards and Adaptive Experience Replay Utilisation in Sparse Reward Environments

open access: yesCAAI Transactions on Intelligence Technology, EarlyView.
ABSTRACT In recent years, the widespread application of deep reinforcement learning (DRL) in autonomous systems has highlighted the importance of achieving high sample efficiency under sparse reward conditions. To improve sample efficiency in sparse reward environments, this paper proposes a reinforcement learning framework built upon the Soft Actor ...
Jingyi Huang   +5 more
wiley   +1 more source

SEMI-MARKOV DECISION PROCESSES WITH COUNTABLE STATE SPACE AND COMPACT ACTION SPACE [PDF]

open access: yes, 1978
We shall be concerned with the optimization problem of semi-Markov decision processes with countable state space and compact action space. Defined is the generalized reward function associated with the semi-Markov decision processes which include the ...
安田, 正實, Yasuda, Masami
core  

Nested Selves: Self‐Organization and Shared Markov Blankets in Prenatal Development in Humans

open access: yesTopics in Cognitive Science, EarlyView., 2023
Abstract The immune system is a central component of organismic function in humans. This paper addresses self‐organization of biological systems in relation to—and nested within—other biological systems in pregnancy. Pregnancy constitutes a fundamental state for human embodiment and a key step in the evolution and conservation of our species. While not
Anna Ciaunica   +3 more
wiley   +1 more source

Alternative seed trait strategies are linked to a trade‐off between spatial and temporal dispersal

open access: yesFunctional Ecology, EarlyView.
Read the free Plain Language Summary for this article on the Journal blog. Abstract Seed dispersal allows plants to seek favourable conditions or spread the risk of unfavourable conditions through space and time. While theory predicts a trade‐off between spatial and temporal dispersal, empirical tests have been stymied by the difficulty of measuring ...
Marina L. LaForgia   +5 more
wiley   +1 more source

The effect of nearby listings on house sale prices in Sydney: A spatio‐temporal regularization approach

open access: yesReal Estate Economics, EarlyView.
Abstract We estimate the price impact of very nearby concurrently listed properties in the Sydney housing market and assess their competition effects. We apply a hedonic model with spatiotemporal effects regularized via a graph Laplacian prior at the month‐by‐SA2 regional level to seven SA4 subregions of metropolitan Sydney. The model structure enables
Willem P. Sijp, Mengheng Li
wiley   +1 more source

Inference on state occupancy in covariate‐driven hidden Markov models

open access: yesMethods in Ecology and Evolution, EarlyView.
Abstract Hidden Markov models (HMMs) are natural and popular tools for analysing animal behaviour based on movement, acceleration and other sensor data. In particular, these models make it possible to infer how the animal's decision‐making process interacts with internal and external drivers by relating the probabilities of switching between distinct ...
Maya Natascha Vienken   +2 more
wiley   +1 more source

A policy gradient method for semi-Markov decision processes [PDF]

open access: yes, 2002
Solving a semi-Markov decision process (SMDP) using value or policy iteration requires precise knowledge of the probabilistic model and suffers from the curse of dimensionality.
Sumeetpal S. Singh   +2 more
core  

Long‐run confidence: Estimating uncertainty when using long‐run multipliers

open access: yesAmerican Journal of Political Science, EarlyView.
Abstract Researchers are often interested in the long‐run relationship (LRR) between variables where the dependent variable has dynamic properties. Though determining the long‐run multiplier (LRM) for an independent variable is straightforward, correctly estimating the significance of the LRM is often difficult, especially when time series are short ...
Mark David Nieman, David A. M. Peterson
wiley   +1 more source

Simulation or cohort models? Continuous time simulation and discretized Markov models to estimate cost-effectiveness [PDF]

open access: yes
The choice of model design for decision analytic models in cost-effectiveness analysis has been the subject of discussion. The current work addresses this issue by noting that, when time is to be explicitly modelled, we need to represent phenomena ...
Marta O Soares, L Canto e Castro
core  

Improving the Understanding of Late Effects of Testicular Cancer in Adolescent and Young Adult Survivors: TRANSCEND‐XR

open access: yesAndrology, EarlyView.
ABSTRACT Background Testicular cancer (TC) is the most common malignancy amongst adolescents and young adults (AYAs) aged 15–39 years assigned male at birth. Survivors often experience late effects of treatment and report unmet supportive care needs.
Mohamad M. Saab   +30 more
wiley   +1 more source

Home - About - Disclaimer - Privacy