← all papers · overview

Video Swin Transformers for Egocentric Video Understanding @ Ego4D Challenges 2022

Abstract

We implemented Video Swin Transformer as a base architecture for the tasks of Point-of-No-Return temporal localization and Object State Change Classification. Our method achieved competitive performance on both challenges.

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).