Results 1 to 10 of about 265,192 (265)
Optimal Policy of Multiplayer Poker via Actor-Critic Reinforcement Learning [PDF]
Poker has been considered a challenging problem in both artificial intelligence and game theory because poker is characterized by imperfect information and uncertainty, which are similar to many realistic problems like auctioning, pricing, cyber security,
Daming Shi +3 more
doaj +2 more sources
A UoI-Optimal Policy for Timely Status Updates with Resource Constraint [PDF]
Timely status updates are critical in remote control systems such as autonomous driving and the industrial Internet of Things, where timeliness requirements are usually context dependent.
Lehan Wang +4 more
doaj +2 more sources
Optimal policy for value-based decision-making [PDF]
Drift diffusion models (DDM) are fundamental to our understanding of perceptual decision-making. Here, the authors show that DDM can implement optimal choice strategies in value-based decisions but require sufficient knowledge of reward contingencies and
Satohiro Tajima +2 more
doaj +2 more sources
Optimal Policies of Social Infrastructures Maintenance using Shock and Damage Model [PDF]
Social infrastructures such as roads and bridges are indispensable for our lives. They have to be maintained continuously and such maintenance has become a big issue in Japan.
Akihiro Yamane +2 more
doaj +1 more source
Optimal Education Plan of Employees Using Maintenance Model [PDF]
The employee education is indispensable for companies to improve productive efficiency and product quality. In general, the employee education is divided into two types, i.e., On-job trainings and Off-job ones, and Off-job trainings are divided into two ...
Shigeshi Yamashita +3 more
doaj +1 more source
We propose an approach for learning optimal tree-based prescription policies directly from data, combining methods for counterfactual estimation from the causal inference literature with recent advances in training globally-optimal decision trees. The resulting method, Optimal Policy Trees, yields interpretable prescription policies, is highly scalable,
Maxime Amram, Jack Dunn, Ying Daisy Zhuo
openaire +3 more sources
Infeasible solutions or negative expected values of future climate information are undesired problems if climate policies are adopted under Cost-Effectiveness Analysis (CEA) to reach uncertain temperature targets.
Mohammad M. Khabbazan
doaj +1 more source
The Bullwhip Effect on the VMI-Supply Chain Management viaSystem Dynamics Approach: The Supply Chain with Two Suppliers and One Retail Channel [PDF]
This work investigates the effect of different inventory policies of a supply chain model using the system dynamics approach which belongs to the class of Vendor Managed Inventory (VMI), automatic pipeline, inventory and order based production control ...
Yahia Zare Mehrjerdi, Alireza Hosseini
doaj +1 more source
Decentralized Policy Optimization
The study of decentralized learning or independent learning in cooperative multi-agent reinforcement learning has a history of decades. Recently empirical studies show that independent PPO (IPPO) can obtain good performance, close to or even better than the methods of centralized training with decentralized execution, in several benchmarks.
Kefan Su, Zongqing Lu 0002
openaire +2 more sources
Competitive Policy Optimization
<p>Submitted - <a href="/records/hzymj-wxd44/files/2006.10611.pdf?download=1">2006.10611.pdf</a></p>
Manish Prajapat +4 more
openaire +4 more sources

