Results 111 to 120 of about 147,564 (269)
Visual teach‐and‐repeat (VTR) navigation allows robots to learn and follow routes without building a full metric map. We show that navigation accuracy for VTR can be improved by integrating a topological map with error‐drift correction based on stereo vision.
Fuhai Ling, Ze Huang, Tony J. Prescott
wiley +1 more source
Continual Learning for Multimodal Data Fusion of a Soft Gripper
Models trained on a single data modality often struggle to generalize when exposed to a different modality. This work introduces a continual learning algorithm capable of incrementally learning different data modalities by leveraging both class‐incremental and domain‐incremental learning scenarios in an artificial environment where labeled data is ...
Nilay Kushawaha, Egidio Falotico
wiley +1 more source
Multimodal Engagement Assessment in Children During Invented Story Paradigm With a Social Robot
A multimodal framework is proposed to assess children's engagement during storytelling interactions with a social robot. Gaze, physiological, and behavioral data are combined and validated against observer ratings. An automated gaze‐labeling strategy is introduced, and supervised classifiers achieve high accuracy. The study supports scalable engagement
Laura Fiorini +7 more
wiley +1 more source
Semantic scene modeling and retrieval
Selected readings in vision and graphics ...
openaire +3 more sources
Multimodal Human–Robot Interaction Using Human Pose Estimation and Local Large Language Models
A multimodal human–robot interaction framework integrates human pose estimation (HPE) and a large language model (LLM) for gesture‐ and voice‐based robot control. Speech‐to‐text (STT) enables voice command interpretation, while a safety‐aware arbitration mechanism prioritizes gesture input for rapid intervention.
Nasiru Aboki +2 more
wiley +1 more source
LLM‐Integrated Human–Robot Interaction System for Microrobots
This paper proposes an LLM‐based control framework for guiding microrobots using human natural language. This framework can convert the natural human speech into safe and executable command sets for reliable navigation in complex environments. The experimental results show high accuracy and robustness in task performance, demonstrating the potential of
Bairong Zhu, Amar Salehi, Tingting Yu
wiley +1 more source
DRIVE‐SAFE evaluates learning‐based, black‐box autonomous driving policies against evolving temporal safety requirements using Signal Temporal Logic robustness metrics. It aggregates distributional robustness measures with domain‐informed weights to guide iterative retraining.
Kristy Sakano +3 more
wiley +1 more source
Learning‐Based Soft Robotic Grasping: Recent Progress and Remaining Challenges
This review analyzes learning‐based soft robotic grasping from a pipeline‐oriented perspective, encompassing soft gripper design, multimodal sensing, and learning‐based planning and control. It surveys key neural network architectures and benchmark datasets and identifies critical challenges such as sim‐to‐real transfer, generalization, and continual ...
Arnab Majumder +3 more
wiley +1 more source
An Integrated NLP‐ML Framework for Property Prediction and Design of Steels
This study presents a data‐driven framework that uses language‐processing techniques to interpret steel processing descriptions and machine‐learning models to predict mechanical properties. By organising complex process histories into meaningful groups and enabling rapid property forecasts, the work supports faster, more informed steel design through ...
Kiran Devraju +5 more
wiley +1 more source
MarginPath is a novel vision‐language system that automates breast cancer margin assessment using a single label‐free multiphoton microscopy image. By integrating tumor‐associated collagen signatures with virtual H&E imaging, it generates accurate margin heatmaps and comprehensive diagnostic reports.
Shu Wang +15 more
wiley +1 more source

