paper-with-me

홈 › Papers

ScaleNet: Scaling up Pretrained Neural Networks with Incremental Parameters

2025-10-21 · Zhiwei Hao, Jianyuan Guo, Li Shen, Kai Han, Yehui Tang, Han Hu, Yunhe Wang arxiv

Recent advancements in vision transformers (ViTs) have demonstrated that larger models often achieve superior performance. However, training these models remains computationally intensive and costly. To address this challenge, we introduce ScaleNet, an efficient approach for scaling ViT models. Unlike conventional training from scratch, ScaleNet facilitates rapid model expansion with negligible increases in parameters, building on existing pretrained models. This offers a cost-effective solution for scaling up ViTs. Specifically, ScaleNet achieves model expansion by inserting additional layers into pretrained ViTs, utilizing layer-wise weight sharing to maintain parameters efficiency. Each added layer shares its parameter tensor with a corresponding layer from the pretrained model. To mitigate potential performance degradation due to shared weights, ScaleNet introduces a small set of adjustment parameters for each layer. These adjustment parameters are implemented through parallel adapter modules, ensuring that each instance of the shared parameter tensor remains distinct and optimized for its specific function. Experiments on the ImageNet-1K dataset demonstrate that ScaleNet enables efficient expansion of ViT models. With a 2$\times$ depth-scaled DeiT-Base model, ScaleNet achieves a 7.42% accuracy improvement over training from scratch while requiring only one-third of the training epochs, highlighting its efficiency in scaling ViTs. Beyond image classification, our method shows significant potential for application in downstream vision areas, as evidenced by the validation in object detection task.

📄 PDF Abstract BibTeX arXiv:2510.18431

Code (0)

등록된 구현이 없습니다.

Tasks

Image ClassificationObject Detection

Results from the Paper

RankTaskDatasetModelMetrics
#1079 Image Classification ImageNet ScaleNet Top 1 Accuracy: 7.42

Similar Papers 제목 키워드 기반

ScaleNet: Searching for the Model to Scale

2022-07-15 · Jiyang Xie, Xiu Su, Shan You, Zhanyu Ma 외

Recently, community has paid increasing attention on model scaling and contributed to developing a model family with a wide spectrum of scales. Current methods either simply resort to a one-shot NAS manner to construct a…

model

ScaleNet: An Unsupervised Representation Learning Method for Limited Information

2023-10-03 · Huili Huang, M. Mahdi Roozbahani

Although large-scale labeled data are essential for deep convolutional neural networks (ConvNets) to learn high-level semantic visual representations, it is time-consuming and impractical to collect and annotate large-sc…

Representation Learning

EcoScaleNet: A Lightweight Multi Kernel Network for Long Sequence 12 lead ECG Classification

2025-10-16 · Dong-Hyeon Kang, Ju-Hyeon Nam, Sang-Chul Lee arxiv

Accurate interpretation of 12 lead electrocardiograms (ECGs) is critical for early detection of cardiac abnormalities, yet manual reading is error prone and existing CNN based classifiers struggle to choose receptive fie…

ECG Classification

ScaleNAS: One-Shot Learning of Scale-Aware Representations for Visual Recognition

2020-11-30 · Hsin-Pai Cheng, Feng Liang, Meng Li, Bowen Cheng 외

Scale variance among different sizes of body parts and objects is a challenging problem for visual recognition tasks. Existing works usually design dedicated backbone or apply Neural architecture Search(NAS) for each tas…

Multi-Person Pose EstimationNeural Architecture SearchOne-Shot LearningPose Estimation+1

ScaleNet: A Shallow Architecture for Scale Estimation

2021-12-09 · CVPR 2022 1 · Axel Barroso-Laguna, Yurun Tian, Krystian Mikolajczyk

In this paper, we address the problem of estimating scale factors between images. We formulate the scale estimation problem as a prediction of a probability distribution over scale factors. We design a new architecture, …

3D ReconstructionCamera Pose EstimationGeometric MatchingPose Estimation