Results 1 to 10 of about 265,192 (265)

Optimal Policy of Multiplayer Poker via Actor-Critic Reinforcement Learning [PDF]

open access: yesEntropy, 2022
Poker has been considered a challenging problem in both artificial intelligence and game theory because poker is characterized by imperfect information and uncertainty, which are similar to many realistic problems like auctioning, pricing, cyber security,
Daming Shi   +3 more
doaj   +2 more sources

A UoI-Optimal Policy for Timely Status Updates with Resource Constraint [PDF]

open access: yesEntropy, 2021
Timely status updates are critical in remote control systems such as autonomous driving and the industrial Internet of Things, where timeliness requirements are usually context dependent.
Lehan Wang   +4 more
doaj   +2 more sources

Optimal policy for value-based decision-making [PDF]

open access: yesNature Communications, 2016
Drift diffusion models (DDM) are fundamental to our understanding of perceptual decision-making. Here, the authors show that DDM can implement optimal choice strategies in value-based decisions but require sufficient knowledge of reward contingencies and
Satohiro Tajima   +2 more
doaj   +2 more sources

Optimal Policies of Social Infrastructures Maintenance using Shock and Damage Model [PDF]

open access: yesInternational Journal of Mathematical, Engineering and Management Sciences, 2021
Social infrastructures such as roads and bridges are indispensable for our lives. They have to be maintained continuously and such maintenance has become a big issue in Japan.
Akihiro Yamane   +2 more
doaj   +1 more source

Optimal Education Plan of Employees Using Maintenance Model [PDF]

open access: yesInternational Journal of Mathematical, Engineering and Management Sciences, 2021
The employee education is indispensable for companies to improve productive efficiency and product quality. In general, the employee education is divided into two types, i.e., On-job trainings and Off-job ones, and Off-job trainings are divided into two ...
Shigeshi Yamashita   +3 more
doaj   +1 more source

Optimal policy trees

open access: yesMachine Learning, 2022
We propose an approach for learning optimal tree-based prescription policies directly from data, combining methods for counterfactual estimation from the causal inference literature with recent advances in training globally-optimal decision trees. The resulting method, Optimal Policy Trees, yields interpretable prescription policies, is highly scalable,
Maxime Amram, Jack Dunn, Ying Daisy Zhuo
openaire   +3 more sources

Cost-Risk Analysis Reconsidered—Value of Information on the Climate Sensitivity in the Integrated Assessment Model PRICE

open access: yesEnergies, 2022
Infeasible solutions or negative expected values of future climate information are undesired problems if climate policies are adopted under Cost-Effectiveness Analysis (CEA) to reach uncertain temperature targets.
Mohammad M. Khabbazan
doaj   +1 more source

The Bullwhip Effect on the VMI-Supply Chain Management viaSystem Dynamics Approach: The Supply Chain with Two Suppliers and One Retail Channel [PDF]

open access: yesInternational Journal of Supply and Operations Management, 2016
This work investigates the effect of different inventory policies of a supply chain model using the system dynamics approach which belongs to the class of Vendor Managed Inventory (VMI), automatic pipeline, inventory and order based production control ...
Yahia Zare Mehrjerdi, Alireza Hosseini
doaj   +1 more source

Decentralized Policy Optimization

open access: yesCoRR, 2022
The study of decentralized learning or independent learning in cooperative multi-agent reinforcement learning has a history of decades. Recently empirical studies show that independent PPO (IPPO) can obtain good performance, close to or even better than the methods of centralized training with decentralized execution, in several benchmarks.
Kefan Su, Zongqing Lu 0002
openaire   +2 more sources

Competitive Policy Optimization

open access: yesCoRR, 2020
<p>Submitted - <a href="/records/hzymj-wxd44/files/2006.10611.pdf?download=1">2006.10611.pdf</a></p>
Manish Prajapat   +4 more
openaire   +4 more sources

Home - About - Disclaimer - Privacy