Results 61 to 70 of about 3,651 (256)
Grounding Large Language Models for Robot Task Planning Using Closed‐Loop State Feedback
BrainBody‐Large Language Model (LLM) introduces a hierarchical, feedback‐driven planning framework where two LLMs coordinate high‐level reasoning and low‐level control for robotic tasks. By grounding decisions in real‐time state feedback, it reduces hallucinations and improves task reliability.
Vineet Bhat +4 more
wiley +1 more source
In this paper, we delve into the innovative application of large language models (LLMs) and their extension, large vision-language models (LVLMs), in the field of remote sensing (RS) image analysis. We particularly emphasize their multi-tasking potential
Yakoub Bazi +4 more
doaj +1 more source
Intelligent Sky Guardians (InSkyGuard) is introduced as a four‐drone swarm that autonomously detects, tracks, and safely captures rogue drones using a coordinated net system. Computer vision and leader–follower control architecture enable synchronized enclosure, while integrated failsafes enhance system reliability. Validated through closed‐environment
Joshua Hastings +6 more
wiley +1 more source
Image captioning is a hot topic that combines a multidiscipline task between computer vision and natural language processing. One of the tasks in the geological field is to make descriptions from the images of geological rocks. The task of a geologist is
Agus Nursikuwagus +3 more
doaj +1 more source
Learning‐Based Soft Robotic Grasping: Recent Progress and Remaining Challenges
This review analyzes learning‐based soft robotic grasping from a pipeline‐oriented perspective, encompassing soft gripper design, multimodal sensing, and learning‐based planning and control. It surveys key neural network architectures and benchmark datasets and identifies critical challenges such as sim‐to‐real transfer, generalization, and continual ...
Arnab Majumder +3 more
wiley +1 more source
Image Description Generation Method by Panoptic Segmentation and Multi-Visual-Feature Fusion [PDF]
Due to their powerful sequence modeling capabilities, Transformer-based image captioning models have demonstrated remarkable performance. However, most of these models typically utilize region visual features to perform encoding and decoding, which ...
LIU Mingming, LU Jinfu, LIU Hao, ZHANG Haiyan
doaj +1 more source
NeuroSuite provides a modular hardware‐software platform integrating Neuroweb and NeuroMaps to enable long‐term, in situ electrophysiological interrogation of air–liquid interface organoid slices while preserving tissue architecture. Its components can be used together or independently to capture real‐time activity, spatial network dynamics, and ...
Belquis Haider +14 more
wiley +1 more source
Subtitling for Deaf and Hard-of-Hearing Children: Not Only Functionality but Art
Subtitling practices for deaf and hard-of-hearing (DHH) children have advanced significantly, addressing both opportunities and challenges in literacy development.
Ewa Domagała-Zyśk
doaj +1 more source
Early Retinal UCHL1 Dysregulation Coupled With Synaptic Loss Reflects Alzheimer's Disease Severity
This study identifies synapse‐enriched deubiquitinase UCHL1 as an early Aβ‐responsive regulator of retinal synaptopathy in Alzheimer's disease. Retinal UCHL1 loss accompanies excitatory synapse degeneration, p75NTR activation, and neuroinflammation, and predicts Braak stage and cognitive decline. Aβ42 fibrils trigger synaptic and UCHL1 depletion before
Altan Rentsendorj +25 more
wiley +1 more source
Can Modern Vision Models Understand the Difference Between an Object and a Look-Alike?
Recent advances in computer vision have yielded models with strong performance on recognition benchmarks; however, significant gaps remain in comparison with human perception.
Itay Cohen, Ethan Fetaya, Amir Rosenfeld
doaj +1 more source

