InstructSee: Instruction-Aware and Feedback-Driven Multimodal Retrieval with Dynamic Query Generation. [PDF]
Gu G, Xue Y, Wu Z, Song L, Liang C.
europepmc +1 more source
Estimating QoE from Encrypted Video Conferencing Traffic. [PDF]
Sidorov M, Birman R, Hadar O, Dvir A.
europepmc +1 more source
End-to-End Teleoperated Driving Video Transmission Under 6G with AI and Blockchain. [PDF]
Frontelo IB +3 more
europepmc +1 more source
A survey on multimodal large language models. [PDF]
Yin S +6 more
europepmc +1 more source
Vision-language models for medical report generation and visual question answering: a review. [PDF]
Hartsock I, Rasool G.
europepmc +1 more source
PlaqueCap: lesion-centered captioning of atherosclerotic plaques in intravascular ultrasound using vision-language models and prompt injection. [PDF]
Ren G +9 more
europepmc +1 more source
Advancing medical imaging with language models: featuring a spotlight on ChatGPT. [PDF]
Hu M +5 more
europepmc +1 more source
Causal Inference Meets Deep Learning: A Comprehensive Survey. [PDF]
Jiao L +9 more
europepmc +1 more source
Quality of Experience (QoE) in Cloud Gaming: A Comparative Analysis of Deep Learning Techniques via Facial Emotions in a Virtual Reality Environment. [PDF]
Jumani AK +6 more
europepmc +1 more source

