Transformers have recently lead to encouraging progress in computer vision. In this work, we present new baselines by improving the original Pyramid Vision Transformer (PVT v1) by adding three designs: (i) a linear complexity attention layer, (ii) an overlapping patch embedding, and (iii) a convolut...
Research Assistant
AI chat, annotations, notes & similar papers
No comments yet
Be the first to share your thoughts!