paper-with-me

DeepViT

2000년 도입 · 논문 2편에서 사용

DeepViT is a type of vision transformer that replaces the self-attention layer within the transformer block with a Re-attention module to address the issue of attention collapse and enables training deeper ViTs.

출처: DeepViT: Towards Deeper Vision Transformer

소개 논문: DeepViT: Towards Deeper Vision Transformer

Image Models · Computer VisionVision Transformers · Computer Vision