Results 91 to 100 of about 6,531,013 (293)

Multiagent Deep Reinforcement Learning Algorithms in StarCraft II: A Review

open access: yesIEEE Access
StarCraft II, as a real-time strategy game, features multiagent collaboration, complex decision-making processes, partially observable environments, and long-term credit assignment; thus, it is an ideal platform for exploring, validating, and optimizing ...
Yanyan Li, Yijun Wang, Yiwei Zhou
doaj   +1 more source

What Do Large Language Models Know About Materials?

open access: yesAdvanced Engineering Materials, EarlyView.
If large language models (LLMs) are to be used inside the material discovery and engineering process, they must be benchmarked for the accurateness of intrinsic material knowledge. The current work introduces 1) a reasoning process through the processing–structure–property–performance chain and 2) a tool for benchmarking knowledge of LLMs concerning ...
Adrian Ehrenhofer   +2 more
wiley   +1 more source

A Survey for Deep Reinforcement Learning Based Network Intrusion Detection

open access: yesApplied AI Letters
Cyber‐attacks are gradually becoming more sophisticated and highly frequent nowadays, and the significance of network intrusion detection systems has become more pronounced.
Wanrong Yang   +3 more
doaj   +1 more source

Improving sample efficiency and exploration in upside-down reinforcement learning

open access: yesJournal of Information and Intelligence
Supervised learning has been demonstrated to be a stable approach for training deep neural networks. Upside-down reinforcement learning solves reinforcement learning problems by using supervised learning, but this method suffers from weak sample ...
Mohammadreza Nakhaei   +1 more
doaj   +1 more source

Thermodynamic Pathways of Nonequilibrium Solidification in Wire‐Arc Additive Manufacturing Fe‐Based Multicomponent Alloy Structures

open access: yesAdvanced Engineering Materials, EarlyView.
Geometry‐driven thermal behavior in wire‐arc additive manufacturing (WAAM) influences microstructural evolution during nonequilibrium solidification of a chemically complex Fe–Cr–Nb–W–Mo–C nanocomposite system. By comparing different deposits configurations, distinct entropy–cooling rate correlations, segregation, and carbide evolution are revealed ...
Blanca Palacios   +5 more
wiley   +1 more source

A Simplified Laminar Flow Model for the Pultrusion of Glass Fiber/Polyethylene Terephthalate Commingled Yarns

open access: yesAdvanced Engineering Materials, EarlyView.
A simplified thermoplastic pultrusion model is developed to predict thermal fields in glass fiber/polyethylene terephthalate (GF/PET) composites with reduced computational cost. By combining effective material homogenization, validation against literature data, and Gaussian‐process‐based optimization, the study reveals how heating limits, pulling speed,
Elder Soares   +3 more
wiley   +1 more source

Learning to Hint for Reinforcement Learning

open access: yesCoRR
Group Relative Policy Optimization (GRPO) is widely used for reinforcement learning with verifiable rewards, but it often suffers from advantage collapse: when all rollouts in a group receive the same reward, the group yields zero relative advantage and thus no learning signal.
Yu Xia 0007   +4 more
openaire   +3 more sources

Delays in Reinforcement Learning

open access: yesCoRR, 2023
Delays are inherent to most dynamical systems. Besides shifting the process in time, they can significantly affect their performance. For this reason, it is usually valuable to study the delay and account for it. Because they are dynamical systems, it is of no surprise that sequential decision-making problems such as Markov decision processes (MDP) can
openaire   +3 more sources

Is an Apple an Orange? A Large Language Model Benchmark for Candidate Term Extraction and Subclass Decisions Against Upper Ontologies in Engineering and Materials Science

open access: yesAdvanced Engineering Materials, EarlyView.
Building machine‐readable vocabularies for materials science is slow, expert‐driven work. This study benchmarks 13 large language models on two of its first steps: finding candidate terms in engineering articles and deciding where they belong in a class hierarchy.
Thomas Bjarsch   +3 more
wiley   +1 more source

A Lightweight Procedural Layer for Hybrid Experimental–Computational Workflows in Materials Science

open access: yesAdvanced Engineering Materials, EarlyView.
We unveil a prototype hybrid‐workflow framework that fuses automatedcomputation with hands‐on experiments. Built atop pyiron, a lightweight, parameterized layer translates procedure descriptions into executable manual steps, syncing instrument settings, human interventions, and data capture in real‐time today.
Steffen Brinckmann   +8 more
wiley   +1 more source

Home - About - Disclaimer - Privacy