Results 31 to 40 of about 1,517,908 (216)
Digitalizing electroplating requires both domain knowledge and interoperability. This work introduces PlatOn, a domain ontology for trivalent chromium plating and coating characterization, and a hybrid pipeline that aligns it to a mid‐level reference ontology by combining eight similarity metrics with language model reasoning. Expert‐validated mappings
Janik Harter +10 more
wiley +1 more source
Speech coding is the most commonly used application of speech processing. Accumulated layers of improvements have however made codecs so complex that optimization of individual modules becomes increasingly difficult. This work introduces machine learning
Bäckström, Tom, Tom Bäckström
core +1 more source
This perspective reframes additive manufacturing for electrical machines as a qualification‐limited materials and architecture design problem. It links process–structure–property–performance relationships to magnetic, conducting, dielectric, and thermal property windows, highlighting where AM can enable segmented magnetic circuits, permanent magnet ...
Dénes Fodor, Loránd Szabó
wiley +1 more source
Conditional emission densities for combining speech enhancement and recognition systems
A novel framework based on conditional emission densities for hidden Markov models (HMMs) is proposed in this contribution to integrate speech enhancement systems with automatic speech recognition systems.
Marc Delcroix +13 more
core +1 more source
Controlled Synthesis of Tri‐ and Multi‐Doped Graphene
This review systematically evaluates synthesis routes for tri‐ and multi‐doped graphene, from hydrothermal and pyrolysis methods to flash Joule heating, critically assessing how each governs dopant incorporation, bonding configuration, and resulting electronic properties.
Maria Hasan +4 more
wiley +1 more source
Wavelets for intonation modeling in HMM speech synthesis [PDF]
The pitch contour in speech contains information about different linguistic units at several distinct temporal scales. At the finest level, the microprosodic cues are purely segmental in nature, whereas in the coarser time scales, lexical tones, word ...
Suni, Antti Santeri +5 more
core
Grounding Large Language Models for Robot Task Planning Using Closed‐Loop State Feedback
BrainBody‐Large Language Model (LLM) introduces a hierarchical, feedback‐driven planning framework where two LLMs coordinate high‐level reasoning and low‐level control for robotic tasks. By grounding decisions in real‐time state feedback, it reduces hallucinations and improves task reliability.
Vineet Bhat +4 more
wiley +1 more source
Speaker adaptation and the evaluation of speaker similarity in the EMIME speech-to-speech translation project [PDF]
This paper provides an overview of speaker adaptation research carried out in the EMIME speech-to-speech translation (S2ST) project. We focus on how speaker adaptation transforms can be learned from speech in one language and applied to the acoustic ...
Tokuda, Keiichi +20 more
core +1 more source
Super-Wideband Spectral Envelope Modeling for Speech Coding
Significant improvements in the quality of speech coders have been achieved by widening the coded frequency range from narrowband to wideband. However, existing speech coders still employ a limited band source-filter model extended by parametric coding ...
Fuchs, Guillaume +5 more
core +1 more source
Continual Learning for Multimodal Data Fusion of a Soft Gripper
Models trained on a single data modality often struggle to generalize when exposed to a different modality. This work introduces a continual learning algorithm capable of incrementally learning different data modalities by leveraging both class‐incremental and domain‐incremental learning scenarios in an artificial environment where labeled data is ...
Nilay Kushawaha, Egidio Falotico
wiley +1 more source

