paper-with-me

홈 › Papers

A Model Zoo of Vision Transformers

2025-04-14 · Damian Falk, Léo Meynent, Florence Pfammatter, Konstantin Schürholt, Damian Borth

The availability of large, structured populations of neural networks - called 'model zoos' - has led to the development of a multitude of downstream tasks ranging from model analysis, to representation learning on model weights or generative modeling of neural network parameters. However, existing model zoos are limited in size and architecture and neglect the transformer, which is among the currently most successful neural network architectures. We address this gap by introducing the first model zoo of vision transformers (ViT). To better represent recent training approaches, we develop a new blueprint for model zoo generation that encompasses both pre-training and fine-tuning steps, and publish 250 unique models. They are carefully generated with a large span of generating factors, and their diversity is validated using a thorough choice of weight-space and behavioral metrics. To further motivate the utility of our proposed dataset, we suggest multiple possible applications grounded in both extensive exploratory experiments and a number of examples from the existing literature. By extending previous lines of similar work, our model zoo allows researchers to push their model population-based methods from the small model regime to state-of-the-art architectures. We make our model zoo available at github.com/ModelZoos/ViTModelZoo.

📄 PDF Abstract BibTeX arXiv:2504.10231

Code (1)

modelzoos/vitmodelzoo 공식 구현 pytorch

Tasks

modelRepresentation Learning

Similar Papers 제목 키워드 기반

A survey of the Vision Transformers and their CNN-Transformer based Variants

2023-05-17 · Asifullah Khan, Zunaira Rauf, Anabia Sohail, Abdul Rehman 외

Vision transformers have become popular as a possible substitute to convolutional neural networks (CNNs) for a variety of computer vision applications. These transformers, with their ability to focus on global relationsh…

Survey

On Convolutional Vision Transformers for Yield Prediction

2024-02-08 · Alvin Inderka, Florian Huber, Volker Steinhage

While a variety of methods offer good yield prediction on histogrammed remote sensing data, vision Transformers are only sparsely represented in the literature. The Convolution vision Transformer (CvT) is being tested to…

Prediction

Vision Transformers: State of the Art and Research Challenges

2022-07-07 · Bo-Kai Ruan, Hong-Han Shuai, Wen-Huang Cheng

Transformers have achieved great success in natural language processing. Due to the powerful capability of self-attention mechanism in transformers, researchers develop the vision transformers for a variety of computer v…

3D ReconstructionImage Segmentationobject-detectionObject Detection+3

AutoTaskFormer: Searching Vision Transformers for Multi-task Learning

2023-04-18 · Yang Liu, Shen Yan, Yuge Zhang, Kan Ren 외

Vision Transformers have shown great performance in single tasks such as classification and segmentation. However, real-world problems are not isolated, which calls for vision transformers that can perform multiple tasks…

Multi-Task LearningNeural Architecture Search

Tensor-to-Image: Image-to-Image Translation with Vision Transformers

2021-10-06 · Yiğit Gündüç

Transformers gain huge attention since they are first introduced and have a wide range of applications. Transformers start to take over all areas of deep learning and the Vision transformers paper also proved that they c…

Image-to-Image TranslationTranslation