Vowel non-vowel based spectral warping and time scale modification for improvement in children’s ASR
Acoustic differences between children’s and adults’ speech causes the degradation in the automatic speech recognition system performance when system trained on adults’ speech and tested on children’s speech. The key acoustic mismatch factors are formant,
Avinash Kumar +5 more
core +1 more source
Speech coding is the most commonly used application of speech processing. Accumulated layers of improvements have however made codecs so complex that optimization of individual modules becomes increasingly difficult. This work introduces machine learning
Bäckström, Tom, Tom Bäckström
core +1 more source
Acoustic analysis and dataset of transitions between coupled rooms
Funding Information: The authors would like to thank the Human Optimised XR (HumOR) Project for funding this research. Publisher Copyright: © 2021 IEEEThe measurement of room acoustics plays a wide role in audio research, from physical acoustics ...
Thomas McKenzie +5 more
core +1 more source
Overlap-add Windows with Maximum Energy Concentration for Speech and Audio Processing
Processing of speech and audio signals with time-frequency representations require windowing methods which allow perfect reconstruction of the original signal and where processing artifacts have a predictable behavior.
Bäckström, T., Tom Backstrom
core +1 more source
Automatic classification of vocal intensity category from speech
Regulation of vocal intensity is a fundamental phenomenon in speech communication. Vocal intensity can be quantified using sound pressure level (SPL), which can be measured easily by recording a standard calibration signal with speech and by comparing ...
Laaksonen, Laura +3 more
core +1 more source
Investigating perceptual discrimination thresholds for attributes of whole-body vibration
Understanding the limitations of haptic perception in humans is critical for the successful design of effective haptic feedback systems, however, it is unclear how perceived discrimination thresholds relate to specific qualitative perceptual attributes ...
Berkay Kullukcu +4 more
doaj +1 more source
Modeling Speech Sound Radiation With Different Degrees of Realism for Articulatory Synthesis
Articulatory synthesis is based on modeling various physical phenomena of speech production, including sound radiation from the mouth. With regard to sound radiation, the most common approach is to approximate it in terms of a simple spherical source of ...
Peter Birkholz +5 more
doaj +1 more source
Comparison of Triply Periodic Minimal Surface Energy Absorbers Under Uniaxial Compressive Loading
This study investigates LCD 3D printed Triply Periodic Minimal Surface (TPMS) structures as mechanical energy absorbers. By comparing various base designs and layered combinations under uniaxial compression, it identifies that a Diamond‐Gyroid sandwich structure offers superior performance.
Sergej Grednev +2 more
wiley +1 more source
Mixture Density Networks, Human Articulatory Data and Acoustic-to-Articulatory Inversion of Continuous Speech [PDF]
Researchers have been investigating methods for retrieving the articulation underlying an acoustic speech signal for more than three decades. A successful method would find many applications, for example: low bit-rate speech coding, helping individuals ...
Richmond, Korin, Richmond, K.; id_orcid
core
A Quantitative Comparison of Epoch Extraction Algorithms for Telephone Speech
Telephone speech is one of the degradations involved in building speech systems in practical environments. The potential use of the speech systems depends on the speech analysis algorithms that can handle different acoustic variations and degradations ...
Kadiri, Sudarsana Reddy +1 more
core +1 more source

