arXiv:2006.11693 Abstract | arXiv Analytics

arXiv:2006.11693 [cs.CV]Abstract References Reviews Resources

Dense-Captioning Events in Videos: SYSU Submission to ActivityNet Challenge 2020

Published 2020-06-21Version 1

This technical report presents a brief description of our submission to the dense video captioning task of ActivityNet Challenge 2020. Our approach follows a two-stage pipeline: first, we extract a set of temporal event proposals; then we propose a multi-event captioning model to capture the event-level temporal relationships and effectively fuse the multi-modal information. Our approach achieves a 9.28 METEOR score on the test set.

Comments: technical report, 4 pages, 2 figures

Categories: cs.CV

Keywords: activitynet challenge, sysu submission, dense-captioning events, dense video captioning task, event-level temporal relationships

Related articles: Most relevant | Search more

arXiv:1705.00754 [cs.CV] (Published 2017-05-02)

Dense-Captioning Events in Videos

Ranjay Krishna, Kenji Hata, Frederic Ren, Li Fei-Fei, Juan Carlos Niebles

arXiv:1710.08011 [cs.CV] (Published 2017-10-22)

ActivityNet Challenge 2017 Summary

Bernard Ghanem et al.

arXiv:1806.04391 [cs.CV] (Published 2018-06-12)

Qiniu Submission to ActivityNet Challenge 2018

Xiaoteng Zhang et al.

arXiv Analytics

arXiv:2006.11693 [cs.CV]Abstract References Reviews Resources

Dense-Captioning Events in Videos: SYSU Submission to ActivityNet Challenge 2020

Links

Toolbox

arXiv:2006.11693 [cs.CV]AbstractReferencesReviewsResources

Dense-Captioning Events in Videos: SYSU Submission to ActivityNet Challenge 2020

Links

Toolbox

arXiv:2006.11693 [cs.CV]Abstract References Reviews Resources