Results 41 to 50 of about 1,921,902 (255)
Multimodal Human–Robot Interaction Using Human Pose Estimation and Local Large Language Models
A multimodal human–robot interaction framework integrates human pose estimation (HPE) and a large language model (LLM) for gesture‐ and voice‐based robot control. Speech‐to‐text (STT) enables voice command interpretation, while a safety‐aware arbitration mechanism prioritizes gesture input for rapid intervention.
Nasiru Aboki +2 more
wiley +1 more source
Adapting WSJ-trained parsers to the British national corpus using in-domain self-training [PDF]
We introduce a set of 1,000 gold standard parse trees for the British National Corpus (BNC) and perform a series of self-training experiments with Charniak and Johnson’s reranking parser and BNC sentences.
Jennifer Foster +7 more
core +2 more sources
LLM‐Integrated Human–Robot Interaction System for Microrobots
This paper proposes an LLM‐based control framework for guiding microrobots using human natural language. This framework can convert the natural human speech into safe and executable command sets for reliable navigation in complex environments. The experimental results show high accuracy and robustness in task performance, demonstrating the potential of
Bairong Zhu, Amar Salehi, Tingting Yu
wiley +1 more source
Ageing is associated with elevated pure-tone thresholds, accompanied by increased difficulties in understanding speech-in-noise. While amplification provides important, but insufficient support, auditory-cognitive training (ACT) might propose a solution.
Vanessa Frei, Nathalie Giroud
doaj +1 more source
Morphemic Structure of Lithuanian Words
The Lithuanian language is a typical flectional language that has a very sophisticated system of grammatical forms and many means of derivation; it is also characterized by uncertain boundaries between morphemes.
Rimkutė Erika +2 more
doaj +1 more source
StackingNet: Collective Inference Across Independent AI Foundation Models
ABSTRACT Artificial intelligence (AI) built on large foundation models has transformed language understanding, computer vision, and reasoning, yet these systems remain isolated and cannot readily share their capabilities. Coordinating the complementary strengths of independently developed, black‐box foundation models is essential for trustworthy ...
Siyang Li +4 more
wiley +1 more source
The multivariate temporal response function (mTRF) is an effective tool for investigating the neural encoding of acoustic and complex linguistic features in natural continuous speech.
Elena Bolt, Nathalie Giroud
doaj +1 more source
IntroductionCooperation is essential to humans and manifests in speech through acoustic convergence, in which interlocutors’ voices become more similar.
Elisa Pellegrino, Volker Dellwo
doaj +1 more source
Neural correlates of linguistic collocations during continuous speech perception
Language is fundamentally predictable, both on a higher schematic level as well as low-level lexical items. Regarding predictability on a lexical level, collocations are frequent co-occurrences of words that are often characterized by high strength of ...
Armine Garibyan +13 more
doaj +1 more source
A unified sequence labeling model for emotion cause pair extraction [PDF]
28th International Conference on Computational Linguistics, December 8-13, 2020, Barcelona, Spain (Online)202402 bcchVersion of ...
Wang, J, Chen, X, Li, Q
core +1 more source

