We study unsupervised video representation learning that seeks to learn both motion and appearance features from unlabeled video only, which can be reused for downstream tasks such as action recognition. This task, however, is extremely challenging due to 1) the highly complex spatial-temporal infor...
Research Assistant
AI chat, annotations, notes & similar papers
No comments yet
Be the first to share your thoughts!