paper-with-me

홈 › Papers

Training independent subnetworks for robust prediction

2020-10-13 · ICLR 2021 1 · Marton Havasi, Rodolphe Jenatton, Stanislav Fort, Jeremiah Zhe Liu, Jasper Snoek, Balaji Lakshminarayanan, Andrew M. Dai, Dustin Tran

Recent approaches to efficiently ensemble neural networks have shown that strong robustness and uncertainty performance can be achieved with a negligible gain in parameters over the original network. However, these methods still require multiple forward passes for prediction, leading to a significant computational cost. In this work, we show a surprising result: the benefits of using multiple predictions can be achieved `for free' under a single model's forward pass. In particular, we show that, using a multi-input multi-output (MIMO) configuration, one can utilize a single model's capacity to train multiple subnetworks that independently learn the task at hand. By ensembling the predictions made by the subnetworks, we improve model robustness without increasing compute. We observe a significant improvement in negative log-likelihood, accuracy, and calibration error on CIFAR10, CIFAR100, ImageNet, and their out-of-distribution variants compared to previous methods.

📄 PDF Abstract BibTeX arXiv:2010.06610

Code (2)

google/uncertainty-baselines 공식 구현 tf
ensta-u2is/torch-uncertainty pytorch

Tasks

Prediction

Similar Papers 제목 키워드 기반

Lottery Pools: Winning More by Interpolating Tickets without Increasing Training or Inference Cost

2022-08-23 · Lu Yin, Shiwei Liu, Meng Fang, Tianjin Huang 외

Lottery tickets (LTs) is able to discover accurate and sparse subnetworks that could be trained in isolation to match the performance of dense networks. Ensemble, in parallel, is one of the oldest time-proven tricks in m…

Cortical representations of Auditory Perception using Graph Independent Component on EEG

2021-10-21 · Pranav Sankhe, Ritik Madan

Recent studies indicate that the neurons involved in a cognitive task aren't locally limited but span out to multiple human brain regions. We obtain network components and their locations for the task of listening to mus…

EEGElectroencephalogram (EEG)

The Championship-Winning Solution for the 5th CLVISION Challenge 2024

2024-06-24 · Sishun Pan, Tingmin Li, Yang Yang

In this paper, we introduce our approach to the 5th CLVision Challenge, which presents distinctive challenges beyond traditional class incremental learning. Unlike standard settings, this competition features the recurre…

Classificationclass-incremental learningClass Incremental LearningContrastive Learning+2

TwIST: Rigging the Lottery in Transformers with Independent Subnetwork Training

2025-11-06 · Michael Menezes, Barbara Su, Xinze Feng, Yehya Farhat 외 arxiv

We introduce TwIST, a distributed training framework for efficient large language model (LLM) sparsification. TwIST trains multiple subnetworks in parallel, periodically aggregates their parameters, and resamples new sub…

Common Complexes of Decompositions and Complex Balanced Equilibria of Chemical Reaction Networks

2021-09-13 · Lauro L. Fontanil, Eduardo R. Mendoza

A decomposition of a chemical reaction network (CRN) is produced by partitioning its set of reactions. The partition induces networks, called subnetworks, that are "smaller" than the given CRN which, at this point, can b…