paper-with-me

홈 › Papers

Revisiting Training-free NAS Metrics: An Efficient Training-based Method

2022-11-16 · Taojiannan Yang, Linjie Yang, Xiaojie Jin, Chen Chen

Recent neural architecture search (NAS) works proposed training-free metrics to rank networks which largely reduced the search cost in NAS. In this paper, we revisit these training-free metrics and find that: (1) the number of parameters (\#Param), which is the most straightforward training-free metric, is overlooked in previous works but is surprisingly effective, (2) recent training-free metrics largely rely on the \#Param information to rank networks. Our experiments show that the performance of recent training-free metrics drops dramatically when the \#Param information is not available. Motivated by these observations, we argue that metrics less correlated with the \#Param are desired to provide additional information for NAS. We propose a light-weight training-based metric which has a weak correlation with the \#Param while achieving better performance than training-free metrics at a lower search cost. Specifically, on DARTS search space, our method completes searching directly on ImageNet in only 2.6 GPU hours and achieves a top-1/top-5 error rate of 24.1\%/7.1\%, which is competitive among state-of-the-art NAS methods. Codes are available at \url{https://github.com/taoyang1122/Revisit_TrainingFree_NAS}

📄 PDF Abstract BibTeX arXiv:2211.08666

Code (1)

taoyang1122/revisit_trainingfree_nas 공식 구현 pytorch

Tasks

GPUNeural Architecture Search

Methods 이 논문이 사용한 방법론

DARTS Differentiable Architecture Search (DART) is a method for efficient architecture search. The search space is made continuous so that the architecture can be optimized with…

Similar Papers 제목 키워드 기반

Revisiting Classifier-Free Guidance Methods in Latent Diffusion Models

2026-08-17 · Artem Sergievskii, Artyom Turevich, Sergey Kastryulin arxiv

Inference-time quality-enhancement methods are an effective and widely adopted means of improving diffusion models without expensive retraining. We study a family of training-free techniques conceptually rooted in Classi…

Semantic correspondence

Revisiting Grammatical Error Correction Evaluation and Beyond

2022-11-03 · Peiyuan Gong, Xuebo Liu, Heyan Huang, Min Zhang

Pretraining-based (PT-based) automatic evaluation metrics (e.g., BERTScore and BARTScore) have been widely used in several sentence generation tasks (e.g., machine translation and text summarization) due to their better …

Grammatical Error CorrectionMachine TranslationSentenceText Summarization

VideoCuRL: Video Curriculum Reinforcement Learning with Orthogonal Difficulty Decomposition

2025-12-31 · Hongbo Jin, Kuanwei Lin, Wenhao Zhang, Yichen Jin 외 arxiv

Reinforcement Learning (RL) is crucial for empowering VideoLLMs with complex spatiotemporal reasoning. However, current RL paradigms predominantly rely on random data shuffling or naive curriculum strategies based on sca…

Reinforcement Learning

Revisiting Data Complexity Metrics Based on Morphology for Overlap and Imbalance: Snapshot, New Overlap Number of Balls Metrics and Singular Problems Prospect

2020-07-15 · José Daniel Pascual-Triana, David Charte, Marta Andrés Arroyo, Alberto Fernández 외

Data Science and Machine Learning have become fundamental assets for companies and research institutions alike. As one of its fields, supervised classification allows for class prediction of new samples, learning from gi…

ClassificationGeneral Classification

Revisiting Gradient Staleness: Evaluating Distance Metrics for Asynchronous Federated Learning Aggregation

2026-03-09 · Patrick Wilhelm, Odej Kao arxiv

In asynchronous federated learning (FL), client devices send updates to a central server at varying times based on their computational speed, often using stale versions of the global model. This staleness can degrade the…

Federated Learning