Results 61 to 70 of about 457 (166)
Abstract Objective The goals of the study were to examine therapists' and clients' emotional states and expressions in an emotionally focused therapy (EFT) couple session, to assess therapeutic attunement between the clients and the therapist, and to explore its alignment with EFT techniques. Background Therapeutic attunement is crucial for fostering a
Gökçenay Başer +3 more
wiley +1 more source
Maleo-Short: An "In-the-Wild" Indonesian Dataset for Speaker Diarization
Speaker diarization (SD), the task of partitioning an audio stream into speaker-homogenous segments, is fundamental for analyzing multi-speaker recordings.
Ardi Mardiana +3 more
doaj +1 more source
Silent Suffering: Using Machine Learning to Measure CEO Depression
ABSTRACT We introduce a novel measure of CEO depression by applying machine learning models that analyze vocal acoustic features from CEOs' conference call recordings. Our research was preregistered via the Journal of Accounting Research's registration‐based editorial process. In this study, we validate this measure and examine associated factors.
SUNG‐YUAN (MARK) CHENG +1 more
wiley +1 more source
ATC-SD Net: Radiotelephone Communications Speaker Diarization Network
This study addresses the challenges that high-noise environments and complex multi-speaker scenarios present in civil aviation radio communications. A novel radiotelephone communications speaker diffraction network is developed specifically for these ...
Weijun Pan +3 more
doaj +1 more source
Abstract INTRODUCTION Digital voice analysis is an emerging tool for differentiating cognitive states, but it poses privacy risks as automated systems may inadvertently identify speakers. METHODS We developed a computational framework to evaluate the trade‐off between voice obfuscation and cognitive assessment accuracy, using pitch‐shifting as a ...
Meysam Ahangaran +5 more
wiley +1 more source
A Script and Tutorial for Using Rev AI's Automatic Speech Transcription
ABSTRACT We introduce Speech Transcriber with Rev AI (STR) ‐ a Python script that allows for easy interfacing with the Rev AI speech transcription service. Recent advancements in technology have led to increased accuracy and affordability of automatic transcription services, making them preferable over the laborious and time‐consuming process of manual
Margaret Broeren +3 more
wiley +1 more source
Abstract Recent advances in generative artificial intelligence (AI) and multimodal learning analytics (MMLA) have allowed for new and creative ways of leveraging AI to support K12 students' collaborative learning in STEM+C domains. To date, there is little evidence of AI methods supporting students' collaboration in complex, open‐ended environments. AI
Clayton Cohn +5 more
wiley +1 more source
Empowering Speaker Segmentation With Self‐Supervised Learning
In this paper, we aim to enhance speaker diarization framework with a novel segmentation model leveraging self‐supervised learning. A lightweight backend for WavLM, comprising a cross‐layer extractor, multi‐head factorized attentive module and a classification head, is proposed to further improve the diarization performance.
Jie Yi, Yunfei Gu, Yinfei Xu
wiley +1 more source
This paper presents the first sentence‐level visual speech recognition (VSR) system specifically designed for the Vietnamese language. We have developed a unique dataset comprising 115 h of video recordings from over 100 speakers, focusing on single‐speaker scenarios.
Phat Nguyen Huu +2 more
wiley +1 more source
Retrieval of TV Talk-Show Speakers by Associating Audio Transcript to Visual Clusters
Retrieval of TV talk-show speakers based on solely visual face recognition is hard because of the significant visual variation caused by illumination, pose, size, and expression, which can exceed those due to identity.
Yina Han, Shanghuan Song, Weikang Zhao
doaj +1 more source

