Results 21 to 30 of about 118,630 (265)

Adaptations in Speech Processing

open access: yes, 2021
Wie sich die Sprachwahrnehmung an ständig eingehende Informationen anpasst, ist eine Schlüsselfrage in der Gedanken- und Gehirnforschung. Die vorliegende Dissertation zielt darauf ab, zum Verständnis von Anpassungen an die Sprecheridentität und Sprachfehler während der Sprachverarbeitung beizutragen und unser Wissen über die Rolle der kognitiven ...
openaire   +2 more sources

NMF-weighted SRP for multi-speaker direction of arrival estimation: robustness to spatial aliasing while exploiting sparsity in the atom-time domain

open access: yesEURASIP Journal on Audio, Speech, and Music Processing, 2021
Localization of multiple speakers using microphone arrays remains a challenging problem, especially in the presence of noise and reverberation. State-of-the-art localization algorithms generally exploit the sparsity of speech in some representation for ...
Sushmita Thakallapalli   +2 more
doaj   +1 more source

Elastic CRFs for Open-Ontology Slot Filling

open access: yesApplied Sciences, 2021
Slot filling is a crucial component in task-oriented dialog systems that is used to parse (user) utterances into semantic concepts called slots. An ontology is defined by the collection of slots and the values that each slot can take.
Yinpei Dai   +5 more
doaj   +1 more source

Federated Learning for privacy-Friendly Health Apps: A Case Study on Ovulation Tracking

open access: yesJournal of Sensor and Actuator Networks
In an era of increasing reliance on digital health solutions, safeguarding user privacy has emerged as a paramount concern. Health applications often need to balance advanced AI functionalities with sufficient privacy measures to ensure user engagement ...
Nikolaos Pavlidis   +12 more
doaj   +1 more source

Coarticulation vs Phonologization: Evidence from Southern Italo-Romance [PDF]

open access: yesJASA Express Letters
This study compares the variety of Italian spoken in Campania and the Italo-Romance variety of Mormanno (Calabria) to investigate the acoustic differences between phonetic vowel-to-vowel coarticulation and phonologized metaphony.
Pia Greca   +2 more
doaj   +1 more source

iCUS: Intelligent CU Size Selection for HEVC Inter Prediction

open access: yesIEEE Access, 2020
The hierarchical quadtree partitioning of Coding Tree Units (CTU) is one of the striking features in HEVC that contributes towards its superior coding performance over its predecessors.
Buddhiprabha Erabadda   +3 more
doaj   +1 more source

Introduction to Digital Speech Processing [PDF]

open access: yesFoundations and Trends® in Signal Processing, 2007
Since even before the time of Alexander Graham Bell’s revolutionary invention, engineers and scientists have studied the phenomenon of speech communication with an eye on creating more efficient and effective systems of human-to-human and human-to-machine communication.
Lawrence R. Rabiner, Ronald W. Schafer
openaire   +1 more source

Effective Exploitation of Posterior Information for Attention-Based Speech Recognition

open access: yesIEEE Access, 2020
End-to-end attention-based modeling is increasingly popular for tackling sequence-to-sequence mapping tasks. Traditional attention mechanisms utilize prior input information to derive attention, which then conditions the output.
Jian Tang   +4 more
doaj   +1 more source

Effective Dereverberation with a Lower Complexity at Presence of the Noise

open access: yesApplied Sciences, 2022
Adaptive beamforming and deconvolution techniques have shown effectiveness for reducing noise and reverberation. The minimum variance distortionless response (MVDR) beamformer is the most widely used for adaptive beamforming, whereas multichannel linear ...
Fengqi Tan, Changchun Bao, Jing Zhou
doaj   +1 more source

Packet Loss Concealment Based on Phase Correction and Deep Neural Network

open access: yesApplied Sciences, 2022
In a packet switching network, the performance of packet loss concealment (PLC) is often affected by inaccurate estimation of phase spectrum of speech signal in the lost packet.
Qiang Ji, Changchun Bao, Zihao Cui
doaj   +1 more source

Home - About - Disclaimer - Privacy