Continual Learning for Multimodal Data Fusion of a Soft Gripper
Models trained on a single data modality often struggle to generalize when exposed to a different modality. This work introduces a continual learning algorithm capable of incrementally learning different data modalities by leveraging both class‐incremental and domain‐incremental learning scenarios in an artificial environment where labeled data is ...
Nilay Kushawaha, Egidio Falotico
wiley +1 more source
CASL-W60: A word-level dataset for central African sign language recognition. [PDF]
Lucky M, Youssouf N, Mahmud H, Hasan MK.
europepmc +1 more source
Auditory–Tactile Congruence for Synthesis of Adaptive Pain Expressions in RoboPatients
In this work, we explore auditory–tactile congruence for synthesizing adaptive vocal pain expressions in robopatients. Using a robopatient platform that integrates vocal pain sounds with palpation forces, we conducted 7680 trials across 20 participants.
Saitarun Nadipineni +4 more
wiley +1 more source
Self-supervised learning with a contrastive VideoMoCo framework for Saudi Arabic sign language recognition using 3D convolutional networks. [PDF]
Rokaya M +4 more
europepmc +1 more source
Multimodal Engagement Assessment in Children During Invented Story Paradigm With a Social Robot
A multimodal framework is proposed to assess children's engagement during storytelling interactions with a social robot. Gaze, physiological, and behavioral data are combined and validated against observer ratings. An automated gaze‐labeling strategy is introduced, and supervised classifiers achieve high accuracy. The study supports scalable engagement
Laura Fiorini +7 more
wiley +1 more source
Bilingual Sign Language Recognition: A YOLOv11-Based Model for Bangla and English Alphabets. [PDF]
Navin N +7 more
europepmc +1 more source
Here, we present a textile, wearable capacitive interface enabling multidirectional remote control by dynamically modulating electrode overlap and spacing via a freely gliding upper electrode. A forearm‐mounted prototype drives robotic and media tasks with 12–15 ms latency, maintains < 0.8% drift after 500 cycles, and remains stably functional at 90 ...
Cagatay Gumus +8 more
wiley +1 more source
Deep computer vision with artificial intelligence based sign language recognition to assist hearing and speech-impaired individuals. [PDF]
Almjally A, Almukadi WS.
europepmc +1 more source
Multimodal Human–Robot Interaction Using Human Pose Estimation and Local Large Language Models
A multimodal human–robot interaction framework integrates human pose estimation (HPE) and a large language model (LLM) for gesture‐ and voice‐based robot control. Speech‐to‐text (STT) enables voice command interpretation, while a safety‐aware arbitration mechanism prioritizes gesture input for rapid intervention.
Nasiru Aboki +2 more
wiley +1 more source
Sign Language Recognition System for Deaf Patients: Protocol for a Systematic Review. [PDF]
Marcolino MS +9 more
europepmc +1 more source

