Results 71 to 80 of about 31,600 (310)

Audio-based event detection for sports video [PDF]

open access: yes, 2003
In this paper, we present an audio-based event detection approach shown to be effective when applied to the Sports broadcast data. The main benefit of this approach is the ability to recognise patterns that indicate high levels of crowd response which ...
Mark Baillie   +3 more
core   +1 more source

LLM‐Integrated Human–Robot Interaction System for Microrobots

open access: yesAdvanced Robotics Research, EarlyView.
This paper proposes an LLM‐based control framework for guiding microrobots using human natural language. This framework can convert the natural human speech into safe and executable command sets for reliable navigation in complex environments. The experimental results show high accuracy and robustness in task performance, demonstrating the potential of
Bairong Zhu, Amar Salehi, Tingting Yu
wiley   +1 more source

Mobile Audiovisual Terminal: System Design and Subjective Testing in DECT and UMTS networks [PDF]

open access: yes, 2000
It is anticipated that there will shortly be a requirement for multimedia terminals that operate via mobile communications systems. This paper presents a functional specification for such a terminal operating at 32 kb/s in a digital European cordless
Pearmain, A, Gill, D, Cosmas, J
core   +1 more source

Moving-Talker, Speaker-Independent Feature Study, and Baseline Results Using the CUAVE Multimodal Speech Corpus

open access: yesEURASIP Journal on Advances in Signal Processing, 2002
Strides in computer technology and the search for deeper, more powerful techniques in signal processing have brought multimodal research to the forefront in recent years.
Patterson Eric K   +3 more
doaj   +1 more source

WS2 Optoelectronic Memristive Reservoir Enabling Ultra‐Low‐Power, Multi‐Task, and Environmentally Stable Neuromorphic Computing

open access: yesAdvanced Science, EarlyView.
WS2‐based in‐memory sensing reservoir computing integrates sensing, memory, and computation in one compact device. It achieves ∼94% N‐MNIST, ∼93% eye motion perception, and ∼89% speech recognition with ultra‐low energy (∼25.5 fJ/spike). The system shows stability at 95% humidity, endurance over 1.5M cycles, and supports synaptic plasticity, enabling ...
Dayanand Kumar   +9 more
wiley   +1 more source

Blind Single Channel Deconvolution using Nonstationary Signal Processing [PDF]

open access: yes, 2003
Blind deconvolution is fundamental in signal processing applications and, in particular, the single channel case remains a challenging and formidable problem. This paper considers single channel blind deconvolution in the case where the degraded observed
Rayner, Peter J. W.   +1 more
core   +1 more source

Adversarial Generation of Voice-Controlled Speaking Face Videos Based on Modal Affine Fusion [PDF]

open access: yesJisuanji gongcheng
Generating speaking face videos from speech, involving the processing both audio and visual modalities, is a current research hotspot. A key challenge is achieving precise alignment between lip movements in the video and the input audio.
CHEN Shihang, SUN Yubao
doaj   +1 more source

Engineering Strategies to Suppress Thermal Runaway Propagation in Lithium‐Ion Battery: Mechanisms, Metrics, Materials, and Evaluation Methods

open access: yesAdvanced Science, EarlyView.
Thermal runaway propagation can transform a single‐cell failure into a system‐level hazard in lithium‐ion battery packs. This review clarifies how heat transfer, gas venting, combustion, and configuration govern cell‐to‐cell failure, and links measurable metrics, pathway‐oriented suppression materials, and experimental/modeling tools to guide safer ...
Jinrong Su   +15 more
wiley   +1 more source

Image volume denoising using a Fourier-wavelet basis [PDF]

open access: yes, 2003
A novel approach to the removal of noise from threedimensional image data is described. The image sequence is represented using a non-adaptive wavelet basis, carefully chosen for its ability to compactly represent locally planar surfaces.
Wilson, Roland   +3 more
core  

Probabilistic Modeling Paradigms for Audio Source Separation [PDF]

open access: yes, 2010
This is the author's final version of the article, first published as E. Vincent, M. G. Jafari, S. A. Abdallah, M. D. Plumbley, M. E. Davies. Probabilistic Modeling Paradigms for Audio Source Separation. In W.
Davies, ME   +14 more
core   +1 more source

Home - About - Disclaimer - Privacy