Results 71 to 80 of about 3,463 (178)
Optoelectronic control of redox‐active polyoxometalate clusters in polymer matrices yields hybrid memristors with switchable volatile and non‐volatile modes, enabling reservoir‐type in‐sensor optical preprocessing and stable multilevel synapses for multimodal neuromorphic computing, including noise‐tolerant audiovisual keyword recognition and hardware ...
Xiangyu Ma +13 more
wiley +1 more source
Mandarin speech‐based early detection of SCD: a feature‐fusion residual network method
Abstract INTRODUCTION Alzheimer's disease (AD) poses a global health challenge. Early intervention during the stage of subjective cognitive decline (SCD) – a potential window for delaying disease progression – is crucial. This study aims to assess an exploratory speech‐based model for rapid SCD screening.
Zhou Liu +6 more
wiley +1 more source
Fusion of Linear and Mel Frequency Cepstral Coefficients for Automatic Classification of Reptiles
Bioacoustic research of reptile calls and vocalizations has been limited due to the general consideration that they are voiceless.
Juan J. Noda +2 more
doaj +1 more source
Flow chart of the base sound trace extraction of speech. ABSTRACT With the acceleration of internationalization, the deficiency of traditional English classroom in English spoken teaching is becoming more obvious, especially the lack of effectiveness and immediate feedback. Therefore, a spoken English assisted training model is proposed.
Lubing Shang, Lu Ma
wiley +1 more source
Automatic Speaker Recognition Based on Mel-Frequency Cepstral Coefficients and Gaussian Mixture Models [PDF]
This paper investigates the task of SR (Speaker Recognition) for the state-of-the-art techniques. The paper initially presents the technical description of automatic SR, followed by the comparative analysis of a number of methods available for feature ...
Sheeraz Memon +2 more
doaj
Automated classification of vowel category and speaker type in the high-frequency spectrum
The high-frequency region of vowel signals (above the third formant or F3) has received little research attention. Recent evidence, however, has documented the perceptual utility of high-frequency information in the speech signal above the traditional ...
Jeremy J. Donai +2 more
doaj +1 more source
A Wireless, Battery‐Free Artificial Throat Patch with Deep Learning for Emotional Speech Recognition
In this work, Xu and co‐workers develop a wireless, battery‐free artificial throat patch system (ATPS) consisting of a carbon nanotube‐based thin‐film strain sensor and a miniaturized flexible printed circuit board, to enable real‐time sensing of throat signals.
Bingxin Xu +10 more
wiley +1 more source
Knocking-sound analysis provides a non-destructive method for assessing durian ripeness; however, most convolutional neural network methods mainly emphasize magnitude information while disregarding phase information that reflects the overall signal ...
Khomdet Phapatanaburi +7 more
doaj +1 more source
Background: Special consideration has recently been given to cepstral analysis with mel-frequency cepstral coefficients (MFCCs). The aim of this study was to assess the applicability of MFCCs in acoustic analysis for diagnosing occupational dysphonia in ...
Ewa Niebudek-Bogusz +3 more
doaj +1 more source
Multilingual Speaker Identification by Combining Evidence from LPR and Multitaper MFCC
In this work, the significance of combining the evidence from multitaper mel-frequency cepstral coefficients (MFCC), linear prediction residual (LPR), and linear prediction residual phase (LPRP) features for multilingual speaker identification with the ...
Nagaraja B.G., Jayanna H.S.
doaj +1 more source

