Results 41 to 50 of about 3,467 (180)

Speech-Based Vehicle Movement Control Solution

open access: yesJournal of Telecommunications and Information Technology, 2021
The article describes a speech-based robotic prototype designed to aid the movement of elderly or handicapped individuals. Mel frequency cepstral coefficients (MFCC) are used for the extraction of speech features and a deep belief network (DBN) is ...
Gurpreet Kaur   +2 more
doaj   +1 more source

Polarization Engineering in Vinylene‐Linked COFs Toward Efficient Neuromorphic Computing

open access: yesAdvanced Science, EarlyView.
A vinyl‐linked covalent organic framework (COF‐DIBPY) was synthesized on a conductive substrate for memristor applications. Redox reactions and bromide migration enable low‐voltage (0.5V) conductance switching with high analog on/off ratio (15.7). The device was further integrated into a speech emotion recognition system for acoustic simulation and ...
Hao Wang   +5 more
wiley   +1 more source

Speech Emotion Recognition Using Cepstral Features Extracted With Gammatone Filter Banks Realized Based on ERB and Mel Frequency Scales

open access: yesIEEE Access
Speech emotion recognition (SER) involves identifying a speaker’s emotional state from their speech utterance. Prior research has explored various cepstral features for developing SER systems. Among these, Mel-frequency cepstral coefficients (MFCC)
Nagarajan Sugan   +2 more
doaj   +1 more source

Mel Frequency Cepstral Coefficient: A Review [PDF]

open access: yesProceedings of the 2nd International Conference on ICT for Digital, Smart, and Sustainable Development, ICIDSSD 2020, 27-28 February 2020, Jamia Hamdard, New Delhi, India, 2021
Shalbbya Ali   +3 more
openaire   +1 more source

Ion‐Gating Reservoir Computing for Preprocessing‐Free Speech Recognition from Throat Vibrations

open access: yesAdvanced Electronic Materials, EarlyView.
This work presents a throat‐mounted mechanoelectric sensor integrated with an ion‐gel/graphene reservoir device for on‐device speech recognition. The system converts raw biomechanical vibrations into rich nonlinear current dynamics, enabling efficient classification through a simple linear readout. The approach highlights a compact and tunable physical‐
Daiki Nishioka   +5 more
wiley   +1 more source

Analysis of the influence of selected audio pre-processing stages on accuracy of speaker language recognition

open access: yesСучасний стан наукових досліджень та технологій в промисловості, 2023
The subject matter of the study is the analysis of the influence of pre-processing stages of the audio on the accuracy of speaker language regognition. The importance of audio pre-processing has grown significantly in recent years due to its key role in
Olesia Barkovska, Anton Havrashenko
doaj   +3 more sources

Hardware Accelerator Design of DCT Algorithm With Unique-Group Cosine Coefficients for Mel-Scale Frequency Cepstral Coefficients

open access: yesIEEE Access, 2022
This study presents a compact $L$ -points discrete cosine transform (DCT) hardware accelerator for $M$ -points Mel-scale Frequency Cepstral Coefficients (MFCC).
Shin-Chi Lai   +6 more
doaj   +1 more source

Soft Active Electromyography Interface for Machine Learning‐Enabled Silent Speech Recognition

open access: yesAdvanced Intelligent Systems, EarlyView.
A soft, hand‐worn electromyography interface enables intent‐driven silent speech recognition without continuous facial attachment. The device integrates liquid‐metal interconnects, a transparent flexible circuit, and elastomer encapsulation with a fingertip electrode that contacts perioral muscles only on demand.
Yuta Kurotaki   +8 more
wiley   +1 more source

Voice authentication module using mel-cepstral coefficients

open access: yesВестник Дагестанского государственного технического университета: Технические науки
Objective. The purpose of the study is to develop and apply a method for extracting information about the identity of users from recordings of their voices using the calculation of mel-cepstral coefficients.Method.
D. A. Elizarov   +2 more
doaj   +1 more source

Passive Acoustic Identification of Social Groups in the Hainan Gibbon

open access: yesRemote Sensing in Ecology and Conservation, EarlyView.
Passive acoustic monitoring offers a non‐invasive means of assessing visually hard‐to‐survey wildlife species with distinctive vocalizations. We evaluated whether deep learning can identify Hainan gibbon (Nomascus hainanus) social groups from their calls.
Emmanuel Kabuga   +14 more
wiley   +1 more source

Home - About - Disclaimer - Privacy