Results 121 to 130 of about 400 (142)
Some of the next articles are maybe not open access.
Multimodal representation fusion method for dense video captioning
Knowledge-Based SystemsYonggang Li
exaly +3 more sources
Attention-based Densely Connected LSTM for Video Captioning
Proceedings of the 27th ACM International Conference on Multimedia, 2019Recurrent Neural Networks (RNNs), especially the Long Short-Term Memory (LSTM), have been widely used for video captioning, since they can cope with the temporal dependencies within both video frames and the corresponding descriptions. However, as the sequence gets longer, it becomes much harder to handle the temporal dependencies within the sequence ...
Yongqing Zhu, Shuqiang Jiang
openaire +1 more source
Event-centric multi-modal fusion method for dense video captioning
Neural Networks, 2022Zhi Chang, Dexin Zhao
exaly
Time–frequency recurrent transformer with diversity constraint for dense video captioning
Information Processing and Management, 2023Ping Li
exaly
MPP-net: Multi-perspective perception network for dense video captioning
Neurocomputing, 2023Shaozu Yuan, Wei Yiwei, Longbiao Wang
exaly
Environment-Aware Dense Video Captioning for IoT-Enabled Edge Cameras
IEEE Internet of Things Journal, 2022Ching-Hu Lu
exaly
Dense Video Captioning With Early Linguistic Information Fusion
IEEE Transactions on Multimedia, 2023Nayyer Aafaq +4 more
openaire +1 more source
Cross-Domain Modality Fusion for Dense Video Captioning
IEEE Transactions on Artificial Intelligence, 2022Nayyer Aafaq +4 more
openaire +2 more sources
Event-Equalized Dense Video Captioning
2025 IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR)Kangyi Wu +7 more
openaire +2 more sources
Egocentric Vehicle Dense Video Captioning
Proceedings of the 32nd ACM International Conference on MultimediaFeiyu Chen 0005 +6 more
openaire +1 more source

