Results 71 to 80 of about 31,600 (310)
Audio-based event detection for sports video [PDF]
In this paper, we present an audio-based event detection approach shown to be effective when applied to the Sports broadcast data. The main benefit of this approach is the ability to recognise patterns that indicate high levels of crowd response which ...
Mark Baillie +3 more
core +1 more source
LLM‐Integrated Human–Robot Interaction System for Microrobots
This paper proposes an LLM‐based control framework for guiding microrobots using human natural language. This framework can convert the natural human speech into safe and executable command sets for reliable navigation in complex environments. The experimental results show high accuracy and robustness in task performance, demonstrating the potential of
Bairong Zhu, Amar Salehi, Tingting Yu
wiley +1 more source
Mobile Audiovisual Terminal: System Design and Subjective Testing in DECT and UMTS networks [PDF]
It is anticipated that there will shortly be a requirement for multimedia terminals that operate via mobile communications systems. This paper presents a functional specification for such a terminal operating at 32 kb/s in a digital European cordless
Pearmain, A, Gill, D, Cosmas, J
core +1 more source
Strides in computer technology and the search for deeper, more powerful techniques in signal processing have brought multimodal research to the forefront in recent years.
Patterson Eric K +3 more
doaj +1 more source
WS2‐based in‐memory sensing reservoir computing integrates sensing, memory, and computation in one compact device. It achieves ∼94% N‐MNIST, ∼93% eye motion perception, and ∼89% speech recognition with ultra‐low energy (∼25.5 fJ/spike). The system shows stability at 95% humidity, endurance over 1.5M cycles, and supports synaptic plasticity, enabling ...
Dayanand Kumar +9 more
wiley +1 more source
Blind Single Channel Deconvolution using Nonstationary Signal Processing [PDF]
Blind deconvolution is fundamental in signal processing applications and, in particular, the single channel case remains a challenging and formidable problem. This paper considers single channel blind deconvolution in the case where the degraded observed
Rayner, Peter J. W. +1 more
core +1 more source
Adversarial Generation of Voice-Controlled Speaking Face Videos Based on Modal Affine Fusion [PDF]
Generating speaking face videos from speech, involving the processing both audio and visual modalities, is a current research hotspot. A key challenge is achieving precise alignment between lip movements in the video and the input audio.
CHEN Shihang, SUN Yubao
doaj +1 more source
Thermal runaway propagation can transform a single‐cell failure into a system‐level hazard in lithium‐ion battery packs. This review clarifies how heat transfer, gas venting, combustion, and configuration govern cell‐to‐cell failure, and links measurable metrics, pathway‐oriented suppression materials, and experimental/modeling tools to guide safer ...
Jinrong Su +15 more
wiley +1 more source
Image volume denoising using a Fourier-wavelet basis [PDF]
A novel approach to the removal of noise from threedimensional image data is described. The image sequence is represented using a non-adaptive wavelet basis, carefully chosen for its ability to compactly represent locally planar surfaces.
Wilson, Roland +3 more
core
Probabilistic Modeling Paradigms for Audio Source Separation [PDF]
This is the author's final version of the article, first published as E. Vincent, M. G. Jafari, S. A. Abdallah, M. D. Plumbley, M. E. Davies. Probabilistic Modeling Paradigms for Audio Source Separation. In W.
Davies, ME +14 more
core +1 more source

