Results 51 to 60 of about 60,353 (264)

MusDiff: A multimodal-guided framework for music generation

open access: yesAlexandria Engineering Journal
Music generation has become a key area in artificial intelligence, achieving significant progress in recent years. However, current research focuses primarily on general music tasks, with limited support for ethnic music. Moreover, the lack of multimodal
Lili Liu, Rui Gong, Yubo Yang
doaj   +1 more source

Principles of libretto translation and problems of multimodal text interpretation

open access: yesRussian Journal of Linguistics, 2022
The multimodal nature of texts to music involves the complex interaction of verbal, auditory, in some cases visual and other components, which determines the functioning of the textual unity.
Аlbina V. Boyarkina
doaj   +1 more source

Artificial Intelligence in Systemic Sclerosis: Clinical Applications, Challenges, and Future Directions

open access: yesArthritis Care &Research, EarlyView.
Systemic sclerosis (SSc) is a rare autoimmune disease defined by immune dysregulation, vasculopathy, and progressive fibrosis of the skin and internal organs. Despite advances in care, major complications such as interstitial lung disease (ILD) and myocardial involvement remain the leading causes of morbidity and mortality.
Cristiana Sieiro Santos   +2 more
wiley   +1 more source

Text-Enhanced Multimodal Method for SAR Ship Classification With Geometry and Polarization Information

open access: yesIEEE Journal of Selected Topics in Applied Earth Observations and Remote Sensing
Synthetic aperture radar (SAR) ship classification is crucial for maritime surveillance. Most existing methods primarily focus on visual or polarimetric features, often constrained by a limited feature set and facing challenges in data diversity and ...
Jinyue Chen   +6 more
doaj   +1 more source

A film adaptation of Stephen King’s novella “Rita Hayworth and Shawshank Redemption” (1982) as an example of a multimodal translation pattern

open access: yesJęzykoznawstwo, 2020
Recently, I came across the statement that „adaptation is a procedure for the translation of the film” (Post). Adaptation reaches not only the increasing circles of the cinema or art, but also our lives, and the very concept of adaptation undergoes ...
Wiktor Szochner
doaj   +1 more source

Multimodal Transformer for Comics Text-Cloze

open access: yes
This work explores a closure task in comics, a medium where visual and textual elements are intricately intertwined. Specifically, Text-cloze refers to the task of selecting the correct text to use in a comic panel, given its neighboring panels. Traditional methods based on recurrent neural networks have struggled with this task due to limited OCR ...
Emanuele Vivoli   +3 more
openaire   +2 more sources

Field Report from Collaborative Research Center 1625: Heterogeneous Research Data Management Using Ontology Representations

open access: yesAdvanced Engineering Materials, EarlyView.
A unified research data management framework for heterogeneous materials data is presented. The system integrates multimodal datasets using ontologies and knowledge graphs, enabling interoperability and FAIR (findable, accessible, interoperable, reusable) data principles. By linking data across scales and workflows, it supports reproducible, Artifitial
Doaa Mohamed   +6 more
wiley   +1 more source

Classification and pragmatic linguistic characteristics of the short vertical video genre in German-speaking educational discourse

open access: yesВестник Самарского университета: История, педагогика, филология
The article considers a new video format that has become extremely popular due to its perfect alignment with mosaic thinking. Although these videos are widely spread on social media, they remain understudied from a linguistic perspective.
M. A. Goncharova
doaj   +1 more source

Is an Apple an Orange? A Large Language Model Benchmark for Candidate Term Extraction and Subclass Decisions Against Upper Ontologies in Engineering and Materials Science

open access: yesAdvanced Engineering Materials, EarlyView.
Building machine‐readable vocabularies for materials science is slow, expert‐driven work. This study benchmarks 13 large language models on two of its first steps: finding candidate terms in engineering articles and deciding where they belong in a class hierarchy.
Thomas Bjarsch   +3 more
wiley   +1 more source

Multimodal foundation models exploit text to make medical image predictions

open access: yesNature Communications
Multimodal foundation models have shown compelling but conflicting performance in medical image interpretation. However, the ways in which these models integrate and prioritize different data modalities, including images and text, remain poorly ...
Thomas A. Buckley   +6 more
doaj   +1 more source

Home - About - Disclaimer - Privacy