Results 31 to 40 of about 49,742 (177)
Retranscrire avec whisper via huma-num
Cela fait maintenant plusieurs mois que le bruit court : un logiciel gratuit permettrait d'obtenir des retranscriptions automatiques d'une qualité excellente. Il s'agit de "Whisper", développé par OpenAI.
Aden Gaide
core +1 more source
HPB SmartNotes: The impact of artificial intelligence on surgeon workload in the outpatient office
In this study, we aimed to evaluate the feasibility, linguistic accuracy, and coherence of medical notes generated by the integration of an automatic speech recognition system (ASR) and a generative pre-trained transformer (GPT) in an outpatient surgical
Rodrigo Antonio Gasque +6 more
doaj +1 more source
Speech Recognition and Synthesis Models and Platforms for the Kazakh Language
With the rapid development of artificial intelligence and machine learning technologies, automatic speech recognition (ASR) and text-to-speech (TTS) have become key components of the digital transformation of society.
Aidana Karibayeva +3 more
doaj +1 more source
AUDIO-TO-TEXT TRANSLATION FOR THE HARD OF HEARING: A WHISPER MODEL-BASED STUDY
This study investigates the effectiveness of the Whisper model for audio-to-text transcription, specifically targeting the enhancement of accessibility for individuals with hearing impairments.
Arypzhan Aben +3 more
doaj +1 more source
Estimación de incertidumbre para un sistema de reconocimiento de voz
Whisper es un sistema de reconocimiento de voz diseñado por la compañía OpenAI, dicho sistema ha sido entrenado con 680,000 horas de datos supervisados multilingües y multitarea recopilados de la web.
Walter Morales-Muñoz +1 more
doaj +1 more source
LLM‐Integrated Human–Robot Interaction System for Microrobots
This paper proposes an LLM‐based control framework for guiding microrobots using human natural language. This framework can convert the natural human speech into safe and executable command sets for reliable navigation in complex environments. The experimental results show high accuracy and robustness in task performance, demonstrating the potential of
Bairong Zhu, Amar Salehi, Tingting Yu
wiley +1 more source
We present a multilingual automatic speech recognition (ASR) system for Kazakh, Russian, and English designed for the trilingual community of Kazakhstan. Although prior research has shown that speech-based text entry can outperform conventional keyboard
Zhanat Makhataeva +5 more
doaj +1 more source
ABSTRACT Aim Children born to mothers with chronic Hepatitis B virus (HBV) infection are at substantial risk of developing chronic HBV‐infection without appropriate perinatal post‐exposure treatment. This study aimed to explore midwives' and public health nurses' (PHNs) experiences with HBV‐post‐exposure treatment for infants and identify factors ...
Brita Askeland Winje +4 more
wiley +1 more source
Whisper is a visually decadent, aurally immersive performance that asks the audience to question ‘what is real’ in a world of increasing technological sophistication. Each audience member is given a set of headphones through which they hear the voices of
Peter Salvatore Petralia (18238750)
core +1 more source
Oral presentations for investor relations, government funding, and technical review are typically prepared without objective, repeatable feedback on delivery.
Yun-Haeng Lee +3 more
doaj +1 more source

