Advances in artificial intelligence: a review for the creative industries. [PDF]
Anantrasirichai N, Zhang F, Bull D.
europepmc +1 more source
A Novel 3D Convolutional Neural Network-Based Deep Learning Model for Spatiotemporal Feature Mapping for Video Analysis: Feasibility Study for Gastrointestinal Endoscopic Video Classification. [PDF]
Dhar MK +14 more
europepmc +1 more source
Parameter efficient multi-model vision assistant for polymer solvation behaviour inference. [PDF]
Liew ZJ, Elkhaiary Z, Lapkin AA.
europepmc +1 more source
Related searches:
Accelerated masked transformer for dense video captioning
Neurocomputing, 2021Abstract Dense video captioning aims to generate dense descriptions for all possible events in an untrimmed video. The task is challenging that it requires accurately localizing events in the video and simultaneously describe each event with a sentence.
Zhou Yu
exaly +3 more sources
Dense Video Captioning Using Graph-Based Sentence Summarization
12 ...
Luping Zhou, Zhiwang Zhang, Wanli Ouyang
exaly +5 more sources
Dense Video Captioning for Incomplete Videos
2021Incomplete video or partially-missing video situations are rarely considered in video captioning research. Previous approaches are mainly trained and evaluated on complete video clip datasets where all the events involved are thoroughly observed. In this work, we formulate the issue of video content description for partially-missing videos.
Xuan Dang +3 more
openaire +1 more source
Hierarchical Language Modeling for Dense Video Captioning
Lecture Notes in Networks and Systems, 2022Padmavathi S
exaly +2 more sources
Event-Centric Hierarchical Representation for Dense Video Captioning
IEEE Transactions on Circuits and Systems for Video Technology, 2021Dense video captioning aims to localize and describe multiple events in untrimmed videos, which is a challenging task that draws attention recently in computer vision. Although existing methods have achieved impressive performance, most of them only focus on local information of event segments or very simple event-level context, overlooking the ...
Teng Wang 0007 +4 more
openaire +2 more sources
Transformer andĀ LLM-Based Captioning Module forĀ Dense Video Captioning
Lecture Notes in Networks and SystemsDvijesh Bhatt
exaly +2 more sources
ADVC: Adversarial dense video captioning with unsupervised pretraining
Image and Vision ComputingJiasi Chen, Wangyu Choi, Jongwon Yoon
exaly +3 more sources

