Results 101 to 110 of about 1,468,270 (262)
Multimodal Feature Learning for Video Captioning [PDF]
Video captioning refers to the task of generating a natural language sentence that explains the content of the input video clips. This study proposes a deep neural network model for effective video captioning.
Incheol Kim, Sujin Lee
core +1 more source
Formation of Distance‐Based Orientation: Political Identity through Relational Positioning in Israel
Distance‐based orientation describes how pejorative labels may serve as anchor points for political identity. Existing research on political labeling has largely emphasized stigmatization, overlooking how labels may acquire durability and orienting capacity without losing pejorative force. Drawing on publicly circulating discourse, we trace positioning
Tammar Friedman, Asaf Saadon
wiley +1 more source
Interactive Videos in Multimodal Listening Assessments: Examining Language Learners' Perspectives
Abstract The academic success of international students who speak English as a second language (L2) hinges on their ability to effectively communicate and comprehend information in English, which requires well‐developed listening skills. Given that real‐world listening mostly involves processing both auditory and visual information, incorporating ...
Shanshan He, Ruslan Suvorov
wiley +1 more source
Video Captioning Using Global-Local Representation. [PDF]
Yan L +6 more
europepmc +1 more source
Abstract Repeated viewing is reportedly a common learning and pedagogical strategy among autonomous second language (L2) learners and language teachers. This experimental study examined the extent to which sequential captioning use facilitates the acquisition of multiword expressions (MWEs) through repeated viewing under incidental learning conditions.
Kenneth W. Y. Li, Yaxin Ni
wiley +1 more source
Video captioning exhibits a complex challenge, particularly due to the increased subject intensity within videos compared to image caption generation. The presence of redundant visual information in video data adds complexity for captioners, making it ...
M. Gowri Shankar, D. Surendran
doaj +1 more source
Semantic-based temporal attention network for Arabic Video Captioning
In recent years, there has been a surge in active research aiming to bridge the gap between computer vision and natural language. In a linguistically diverse region like the Arab world, it is essential to establish a mechanism that facilitates the ...
Adel Jalal Yousif, Mohammed H. Al-Jammas
doaj +1 more source
Research on Video Captioning Based on Multifeature Fusion. [PDF]
Zhao H, Guo L, Chen Z, Zheng H.
europepmc +1 more source
Video Captioning via Relation-Aware Graph Learning
Video Captioning via Relation-Aware Graph ...
T Zhang (5107766) +6 more
core

