Results 41 to 50 of about 212,805 (267)
Ramp Metering Control Based on the Q-Learning Algorithm
Modern urban highways are under the influence of increased traffic demand and cannot fulfill the desired level of service anymore. In most of the cases there is no space available for any infrastructure building.
Ivanjko Edouard +5 more
doaj +1 more source
Q-learning is widely used algorithm in reinforcement learning community. Under the lookup table setting, its convergence is well established. However, its behavior is known to be unstable with the linear function approximation case. This paper develops a new Q-learning algorithm that converges when linear function approximation is used.
Han-Dong Lim, Donghwan Lee 0002
openaire +3 more sources
Fragile X syndrome (FXS) is a neurodevelopmental disorder caused by hypermethylation of expanded CGG repeats (>200) in the FMR1 gene leading to gene silencing and loss of Fragile X Messenger Ribonucleoprotein (FMRP) expression. FMRP plays important roles
James J. Fink +20 more
doaj +1 more source
Complexification through gradual involvement and reward Providing in deep reinforcement learning
Training a relatively big neural network within the framework of deep reinforcement learning that has enough capacity for complex tasks is challenging. In real life the process of task solving requires system of knowledge, where more complex skills are ...
E. V. Rulko,
doaj +1 more source
Q-learning (Watkins, 1989) is a simple way for agents to learn how to act optimally in controlled Markovian domains. It amounts to an incremental method for dynamic programming which imposes limited computational demands. It works by successively improving its evaluations of the quality of particular actions at particular states.
Watkins, C., Dayan, P.
openaire +2 more sources
Smooth Q-learning: Accelerate Convergence of Q-learning Using Similarity
An improvement of Q-learning is proposed in this paper. It is different from classic Q-learning in that the similarity between different states and actions is considered in the proposed method. During the training, a new updating mechanism is used, in which the Q value of the similar state-action pairs are updated synchronously. The proposed method can
Wei Liao, Xiaohui Wei, Jizhou Lai
openaire +2 more sources
Modulation of Homer1 EVH1 domain internal dynamics by putative autism‐associated mutations
The putative autism‐associated M65I and S97L variants of the EVH1 domain of the postsynaptic scaffold protein Homer1 do not exhibit substantial changes in their overall structure or partner binding. Both of them, but especially the M65I variant, show altered internal dynamics relative to the wild‐type domain on the μs‐ms timescale, indicated by the ...
Fanni Farkas +6 more
wiley +1 more source
Offloading decision algorithm based on reinforcement learning for mobile edge computing
For the problem of computing offloading decision in mobile edge computing, this paper proposes an offloading decision algorithm based on enhanced learning in multiuser MEC system.
Yang Ge, Zhang Heng
doaj +1 more source
Identification of the plant mitochondrial OrfX protein: A mass spectrometry approach
The mitochondrial genome of plants contains an open reading frame, orfx, which encodes a rare protein that has so far escaped mass spectrometric detection. The protein resembles the c‐subunit of bacterial twin‐arginine‐motif‐dependent protein translocases (TatC).
Matthias Döring +3 more
wiley +1 more source
This study shows that lung adenocarcinomas exploit developmental branching morphogenesis to acquire a therapy resistant basal‐like tumour cell state. This process was found to be regulated by combined TP53 loss‐of‐function and type‐I interferon signalling, identifying a novel axis for biomarker and therapeutic target discovery.
Kamila J Bienkowska +13 more
wiley +1 more source

