Results 41 to 50 of about 3,651 (256)
Methodologies that utilize Deep Learning offer great potential for applications that automatically attempt to generate captions or descriptions about images and video frames.
Soheyla Amirian +3 more
doaj +1 more source
Magnetic tunnel junctions (MTJs) using MgO tunnel barriers face challenges of high resistance‐area product and low tunnel magnetoresistance (TMR). To discover alternative materials, Literature Enhanced Ab initio Discovery (LEAD) is developed. The LEAD‐predicted materials are theoretically evaluated, showing that MTJs with dusting of ScN or TiN on ...
Sabiq Islam +6 more
wiley +1 more source
Offline visual aid system for the blind based on image captioning
In view of the inconveniences of existing visual aid systems for the blind, the method of running the image captioning model on portable mobile devices based on model pruning was discussed.Model pruning techniques and image captioning models were ...
Yue CHEN +3 more
doaj +2 more sources
Can Audio Captions Be Evaluated With Image Caption Metrics?
ICASSP ...
Zelin Zhou +5 more
openaire +4 more sources
A wearable electrochemical microneedle patch integrates Prussian Blue redox transduction with molecularly imprinted polymer recognition for reagent‐free sampling and detection of thrombo‐inflammatory biomarkers in dermal interstitial fluid. The platform tracks thrombin and inflammatory cytokines with sensitive in vitro, ex vivo, and in vivo performance,
Mahmoud Ayman Saleh +11 more
wiley +1 more source
Abstract Dense video captioning involves detecting and describing events within video sequences. Traditional methods operate in an offline setting, assuming the entire video is available for analysis. In contrast, in this work we introduce a groundbreaking paradigm: Live Video Captioning (LVC), where captions must
Blanco Fernández, Eduardo +4 more
openaire +6 more sources
The energetic offset between the donor and the acceptor components in organic photoactive layers is central to the tradeoff between photovoltage and photocurrent losses. This Perspective covers the most important issues surrounding this topic in non‐fullerene acceptor blends, from the difficulty of accurately determining state energies and driving ...
Dieter Neher, Manasi Pranav
wiley +1 more source
AVCaps: An Audio-Visual Dataset With Modality-Specific Captions
This paper introduces AVCaps, an audio-visual dataset that contains separate textual captions for the audio, visual, and audio-visual contents of video clips. The dataset contains 2061 video clips constituting a total of 28.8 hours.
Parthasaarathy Sudarsanam +3 more
doaj +1 more source
A novel engineering strategy that establishes material design principles for incorporating anti‐inflammatory steroids into lipid nanoparticles to reduce LNP‐induced inflammation while retaining mRNA delivery. Results are validated in vitro and in three animal models of inflammation and autoimmunity in vivo.
Ajay S. Thatte +21 more
wiley +1 more source
Aerial Image Analysis: When LLMs Assist (And When Not)
Large language models (LLMs) have shown remarkable results when tasked with the analysis and production of texts or images and for captioning images.
Salvatore Calcagno +3 more
doaj +1 more source

