This notebook paper presents an overview and comparative analysis of our systems designed for the following three tasks in ActivityNet Challenge 2019: trimmed action recognition, dense-captioning events in videos, and spatio-temporal action localization.
Related papers
Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).