SPARK: sparse-perception action recognition with keyframes for quadruped robots. [PDF]
Park S, Choi AJ.
europepmc +1 more source
MIRA-CAP: Memory-Integrated Retrieval-Augmented Captioning for State-of-the-Art Image and Video Captioning. [PDF]
Umirzakova S +4 more
europepmc +1 more source
AI-powered BlindSpot VisionGuide system on raspberry Pi for enhancing independence of visually impaired users. [PDF]
Sudha M +3 more
europepmc +1 more source
An innovative multi-head attention mechanism-driven recurrent neural network model with feature representation fusion for enhanced image captioning to assist individuals with visual impairments. [PDF]
Asiri MM +3 more
europepmc +1 more source
A Systematic Review of Deep Learning and Machine Learning Applications in Longitudinal Multimodal Clinical Data. [PDF]
Pan J +11 more
europepmc +1 more source
Lightweight visual accessibility LLaVA architecture. [PDF]
Han Z, Liu X, Hao J.
europepmc +1 more source
Class-dependent and cross-modal memory network considering sentimental features for video-based captioning. [PDF]
Xiong H, Zhou Y, Liu J, Cai Y.
europepmc +1 more source
Inclusive instructional communication in higher education: A comparative policy analysis of South Africa and Malawi. [PDF]
De Souza B.
europepmc +1 more source
EdgeVidCap: A Channel-Spatial Dual-Branch Lightweight Video Captioning Model for IoT Edge Cameras. [PDF]
Guo L +9 more
europepmc +1 more source
MSSA: memory-driven and simplified scaled attention for enhanced image captioning. [PDF]
Hossain MA +5 more
europepmc +1 more source

