Visual Robustness Benchmark for Visual Question Answering (VQA)
Can Visual Question Answering (VQA) systems perform just as well when deployed in the real world? Or are they susceptible to realistic corruption effects e.g. image blur, which can be detrimental in sensitive applications, such as medical VQA?
Tashdeed, Ishmam +5 more
core
Intravenous extravasation report generation using deep learning, generative artificial intelligence, and visual question answering techniques. [PDF]
Munthuli A +7 more
europepmc +1 more source
PeFoMed: Parameter efficient fine-tuning of multimodal large language models for medical CXR. [PDF]
Liu G +5 more
europepmc +1 more source
Agentic systems in computational pathology: architectures, evidence, and translational challenges. [PDF]
Lu X +13 more
europepmc +1 more source
LMM-VQA: Advancing Video Quality Assessment with Large Multimodal Models
The explosive growth of videos on streaming media platforms has underscored the urgent need for effective video quality assessment (VQA) algorithms to monitor and perceptually optimize the quality of streaming videos.
Sun, Wei +8 more
core
The New Frontier of Quality Evaluation for Visual Sensors: A Survey of Large Multimodal Model-Based Methods. [PDF]
Ge Q, Min X, Wu S, Li Y, Zhai G.
europepmc +1 more source
Research on the application of LLaVA model based on QLoRA fine-tuning in medical teaching. [PDF]
Zhou S, Qin F.
europepmc +1 more source
ViDscribe: Multimodal AI for Customizing Audio Description and Question Answering in Online Videos. [PDF]
Cheema MS +3 more
europepmc +1 more source
D<sup>2</sup>MNet: Difference-Aware Decoupling and Multi-Prompt Learning for Medical Difference Visual Question Answering. [PDF]
Lai L, Ou W, Gou J, Liu Z.
europepmc +1 more source
Contrastive learning-based video quality assessment-jointed video vision transformer for video recognition. [PDF]
Sun J, Mahoor M.
europepmc +1 more source

