In this paper, we question if self-supervised learning provides new properties to Vision Transformer (ViT) [16] that stand out compared to convolutional networks (convnets). Beyond the fact that adapting self-supervised methods to this architecture works particularly well, we make the following obse...
Research Assistant
AI chat, annotations, notes & similar papers
No comments yet
Be the first to share your thoughts!