Results 81 to 90 of about 49,742 (177)
Effectiveness of Whisper's Fine-Tuning for Domain-Specific Use Cases in the Industry
: The integration of Speech-to-Text (STT) technology has the potential to enhance the efficiency of industrial workflows. However, standard speech models demonstrate suboptimal performance in domain-specific use cases.
Daniel Pawlowicz +2 more
semanticscholar +1 more source
As the line between human speakers and “AI-generated” voices becomes increasingly blurred, it is important to understand how sociolinguistic knowledge affects human-computer interaction. Human listeners have been shown to rely on real-world biases, along
Eve Fleisig +6 more
semanticscholar +1 more source
Whisper in Medusa\u27s Ear: Multi-head Efficient Decoding for Transformer-based ASR [PDF]
Large transformer-based models have significant potential for speech transcription and translation. Their self-attention mechanisms and parallel processing enable them to capture complex patterns and dependencies in audio sequences.
Keshet, Joseph +4 more
core +1 more source
Riconoscimento del parlato mediante OpenAI Whisper
Questa tesi si propone di implementare e analizzare un sistema di riconoscimento vocale in tempo reale in locale utilizzando OpenAI Whisper, un modello avanzato basato su tecniche di deep learning. Whisper rappresenta lo stato dell’arte nella comprensione del parlato umano e si distingue per essere un modello open source.
openaire +1 more source
Analysis and Categorization of the Rusyn Language Using the Whisper Model: Demographic Influences on Linguistic Convergence [PDF]
The article presents a detailed linguistic analysis of the Rusyn language, focusing on its complex and evolving features, such as pronunciation, as well as individual, regional, and historical variabilities.
Małecki, Paweł
core +1 more source
mmWave-Whisper: Phone Call Eavesdropping and Transcription Using Millimeter-Wave Radar [PDF]
This paper introduces mmWave-Whisper, a system that demonstrates the feasibility of full-corpus automated speech recognition (ASR) on phone calls eavesdropped remotely using off-the-shelf frequency modulated continuous wave (FMCW) millimeter-wave radars.
Basak, Suryoday +2 more
core +1 more source
Whisper Finetuning on Nepali Language
Despite the growing advancements in Automatic Speech Recognition (ASR) models, the development of robust models for underrepresented languages, such as Nepali, remains a challenge.
Rijal, Sanjay +4 more
core
Simul-Whisper: Attention-Guided Streaming Whisper with Truncation Detection [PDF]
As a robust and large-scale multilingual speech recognition model, Whisper has demonstrated impressive results in many low-resource and out-of-distribution scenarios.
Wang, Haoyu +4 more
core +1 more source

