NanoNet: Parameter-Efficient Learning with Label-Scarce Supervision for Lightweight Text Mining Model
The lightweight semi-supervised learning (LSL) strategy provides an effective approach of conserving labeled samples and minimizing model inference costs. Prior research has effectively applied knowledge transfer learning and co-training regularization from large to small models in LSL. However, such training strategies are computationally intensive and prone to local optima, thereby increasing the difficulty of finding the optimal solution. This has prompted us to investigate the feasibility of integrating three low-cost scenarios for text mining tasks: limited labeled supervision, lightweight fine-tuning, and rapid-inference small models. We propose NanoNet, a novel framework for lightweight text mining that implements parameter-efficient learning with limited supervision. It employs online knowledge distillation to generate multiple small models and enhances their performance through mutual learning regularization. The entire process leverages parameter-efficient learning, reducing training costs and minimizing supervision requirements, ultimately yielding a lightweight model for downstream inference.
Code (0)
등록된 구현이 없습니다.
Tasks
Knowledge DistillationTransfer LearningSimilar Papers 제목 키워드 기반
NanoNet: Real-Time Polyp Segmentation in Video Capsule Endoscopy and Colonoscopy
Deep learning in gastrointestinal endoscopy can assist to improve clinical performance and be helpful to assess lesions more accurately. To this extent, semantic segmentation methods that can perform automated real-time …
Colorectal Polyps CharacterizationInstrument RecognitionMedical Image SegmentationReal-Time Semantic Segmentation+3Rethinking Remaining Useful Life Prediction with Scarce Time Series Data: Regression under Indirect Supervision
Supervised time series prediction relies on directly measured target variables, but real-world use cases such as predicting remaining useful life (RUL) involve indirect supervision, where the target variable is labeled a…
PredictionregressionTime SeriesTime Series Prediction+1A Comprehensive Survey on Hybrid Communication for Internet of Nano-Things in Context of Body-Centric Communications
With the huge advancement of nanotechnology over the past years, the devices are shrinking into micro-scale, even nano-scale. Additionally, the Internet of nano-things (IoNTs) are generally regarded as the ultimate forma…
SCALER: SAM-Enhanced Collaborative Learning for Label-Deficient Concealed Object Segmentation
Existing methods for label-deficient concealed object segmentation (LDCOS) either rely on consistency constraints or Segment Anything Model (SAM)-based pseudo-labeling. However, their performance remains limited due to t…
Object SegmentationWho Needs Labels? Adapting Vision Foundation Models With the Metadata You Already Have
We propose a label-free approach to adapt powerful but generic vision foundation models to specialized scientific domains. Standard supervised fine-tuning is often ill-suited to these settings: labels are scarce, and tas…
Unsupervised Domain Adaptation