Results 21 to 30 of about 49,742 (177)
Classification of heart sound signals with Whisper model
Heart sounds, or phonocardiograms (PCG), are important for diagnosing cardiovascular conditions, providing a non-invasive means to assess heart function through auscultation.
Yakoub Bazi, Nassim Ammour
exaly +2 more sources
Transcription of Audio and Video with OpenAI’s Whisper
Margarete Wilkey, Geoffrey Wood
openaire +2 more sources
Whisper Game: Practicing Attention Through Caring and Pacing [PDF]
In “Whisper Game: Practicing Attention Through Caring and Pacing,” Basia Sliwinska and Caroline Stevenson experienced embodied knowledge in conversation with Judah Attille, Giovanna Bragaglia, Maria Costantino, Fabiola Fiocco and Alison Green by playing ...
Stevenson, Caroline, Sliwinska, Basia
core +5 more sources
Automated Caption Generation for Video Call with Language Translation [PDF]
In the modern era, virtual communication between individuals is common. Many people’s lives have been made simpler in a number of circumstances by providing subtitles, generating automated captions for social media videos, and language translation from a
Polepaka Sanjeeva +4 more
doaj +1 more source
AI-Powered Audio Summarization and Ethical Content Analysis Using OpenAI Whisper
Samyak Vamshi Shaga
openaire +2 more sources
Adapting OpenAI’s Whisper for Speech Recognition on Code-Switch Mandarin-English SEAME and ASRU2019 Datasets [PDF]
This paper reports on SOTA results achieved using openAI’s Whisper model with adaptation on different adaptation corpus sizes for two established code-switch Mandarin/English corpus - namely SEAME and ASRU2019 corpora.Two key experiments were conducted ...
Yuhang Yang +4 more
semanticscholar +1 more source
BENCHMARKING WHISPER OPENAI ON SARAWAK LANGUAGES [PDF]
The end-to-end (E2E) model is influentially reshaping the automatic speech recognition (ASR) scene, supplanting traditional ASR models such as the Hidden Markov model (HMM) and Deep Neural Network (DNN)-based hybrid models.
GERALD EINSTEIN CORNELIUS
core
BackgroundAutomatic speech recognition (ASR) technology is increasingly being used for transcription in clinical contexts. Although there are numerous transcription services using ASR, few studies have compared the word error rate
Salman Seyedi +11 more
doaj +1 more source
Reconstruction of Phonated Speech from Whispers Using Formant-Derived Plausible Pitch Modulation [PDF]
Whispering is a natural, unphonated, secondary aspect of speech communications for most people. However, it is the primary mechanism of communications for some speakers who have impaired voice production mechanisms, such as partial laryngectomees, as ...
Sharifzadeh, Hamid Reza +4 more
core +1 more source
Whisper-to-speech conversion using restricted Boltzmann machine arrays [PDF]
Whispers are a natural vocal communication mechanism, in which vocal cords do not vibrate normally. Lack of glottal-induced pitch leads to low energy, and an inherent noise-like spectral distribution reduces intelligibility.
Li, Jing-jie +3 more
core +1 more source

