Results 51 to 60 of about 54,505 (261)
High Capacity Audio Steganography Based on Contourlet Transform
The science of hiding information behind other cover media file is known steganography. Audio steganography means that a secret message is hidden by embedding it in an audio file.
Abbas Salman Hameed
doaj
Continual Learning for Multimodal Data Fusion of a Soft Gripper
Models trained on a single data modality often struggle to generalize when exposed to a different modality. This work introduces a continual learning algorithm capable of incrementally learning different data modalities by leveraging both class‐incremental and domain‐incremental learning scenarios in an artificial environment where labeled data is ...
Nilay Kushawaha, Egidio Falotico
wiley +1 more source
Audio watermarking: a way to stationnarize audio signals [PDF]
Audio watermarking is usually used as a multimedia copyright protection tool or as a system that embed metadata in audio signals. In this paper, watermarking is viewed as a preprocessing step for further audio processing systems: the watermark signal conveys no information, it is rather used to modify the statistical characteristics of an audio signal,
Djaziri-Larbi, Sonia, Jaidane, Mériem
openaire +2 more sources
Auditory–Tactile Congruence for Synthesis of Adaptive Pain Expressions in RoboPatients
In this work, we explore auditory–tactile congruence for synthesizing adaptive vocal pain expressions in robopatients. Using a robopatient platform that integrates vocal pain sounds with palpation forces, we conducted 7680 trials across 20 participants.
Saitarun Nadipineni +4 more
wiley +1 more source
Multimodal Engagement Assessment in Children During Invented Story Paradigm With a Social Robot
A multimodal framework is proposed to assess children's engagement during storytelling interactions with a social robot. Gaze, physiological, and behavioral data are combined and validated against observer ratings. An automated gaze‐labeling strategy is introduced, and supervised classifiers achieve high accuracy. The study supports scalable engagement
Laura Fiorini +7 more
wiley +1 more source
Multimodal Human–Robot Interaction Using Human Pose Estimation and Local Large Language Models
A multimodal human–robot interaction framework integrates human pose estimation (HPE) and a large language model (LLM) for gesture‐ and voice‐based robot control. Speech‐to‐text (STT) enables voice command interpretation, while a safety‐aware arbitration mechanism prioritizes gesture input for rapid intervention.
Nasiru Aboki +2 more
wiley +1 more source
LLM‐Integrated Human–Robot Interaction System for Microrobots
This paper proposes an LLM‐based control framework for guiding microrobots using human natural language. This framework can convert the natural human speech into safe and executable command sets for reliable navigation in complex environments. The experimental results show high accuracy and robustness in task performance, demonstrating the potential of
Bairong Zhu, Amar Salehi, Tingting Yu
wiley +1 more source
Objectives. The aim of this study is to develop and analyze parameters for a multifunctional audio module based on the ADAU1701 audio digital signal processor in the SigmaStudio environment.
A. V. Gevorsky +2 more
doaj +1 more source
Special Issue on “Sound and Music Computing”
Sound and music computing is a young and highly multidisciplinary research field. [...]
Tapio Lokki +3 more
doaj +1 more source
Compressive Sampling with Multiple Bit Spread Spectrum-Based Data Hiding
We propose a novel data hiding method in an audio host with a compressive sampling technique. An over-complete dictionary represents a group of watermarks.
Gelar Budiman +2 more
doaj +1 more source

