paper-with-me

홈 › Papers

InTrain: Intrinsic Trainability for Zero-Cost Neural Architecture Search

2026-06-17 · Qinqin Zhou, Fuhai Chen, Jipeng Wu, Zhiwei Chen, Zhikai Hu, Weiwei Cai arxiv

Training-free neural architecture search promises efficient discovery of high-performance networks without costly training. However, existing zero-cost proxies rely on fragmented heuristics that fail to capture the fundamental question: what makes an architecture trainable? This paper introduces Intrinsic Trainability (InTrain), a unified theoretical proxy that formalizes trainability as an architectural invariant emerging from two synergistic components: geometric capacity and optimization resilience. We operationalize intrinsic trainability through analysis of neural information processing. Geometric capacity is quantified via the participation ratio of activation covariance eigenspectrum, capturing the effective dimensionality of representation manifolds. Optimization resilience is measured through cumulative gradient health, assessing the robustness of backpropagation across network depth. InTrain synthesizes these dimensions through a scale-invariant multiplicative coupling, which we hypothesize is essential for capturing their synergistic, non-additive relationship. Extensive experiments on standard NAS benchmarks and search spaces demonstrate that InTrain achieves ranking correlations on par with state-of-the-art ensemble-based proxies and outperforms other single-metric methods.

📄 PDF Abstract BibTeX arXiv:2606.18676

Code (0)

등록된 구현이 없습니다.

Tasks

Neural Architecture Search

Similar Papers 제목 키워드 기반

Can Error Mitigation Improve Trainability of Noisy Variational Quantum Algorithms?

2021-09-02 · Samson Wang, Piotr Czarnik, Andrew Arrasmith, M. Cerezo 외

Variational Quantum Algorithms (VQAs) are often viewed as the best hope for near-term quantum advantage. However, recent studies have shown that noise can severely limit the trainability of VQAs, e.g., by exponentially f…

regression

AZ-NAS: Assembling Zero-Cost Proxies for Network Architecture Search

2024-03-28 · CVPR 2024 1 · Junghyup Lee, Bumsub Ham

Training-free network architecture search (NAS) aims to discover high-performing networks with zero-cost proxies, capturing network characteristics related to the final performance. However, network rankings estimated by…

Reinfier and Reintrainer: Verification and Interpretation-Driven Safe Deep Reinforcement Learning Frameworks

2024-10-19 · Zixuan Yang, Jiaqi Zheng, Guihai Chen

Ensuring verifiable and interpretable safety of deep reinforcement learning (DRL) is crucial for its deployment in real-world applications. Existing approaches like verification-in-the-loop training, however, face challe…

Deep Reinforcement Learning

Trainability of Dissipative Perceptron-Based Quantum Neural Networks

2020-05-26 · Kunal Sharma, M. Cerezo, Lukasz Cincio, Patrick J. Coles

Several architectures have been proposed for quantum neural networks (QNNs), with the goal of efficiently performing machine learning tasks on quantum data. Rigorous scaling results are urgently needed for specific QNN c…

When the Left Foot Leads to the Right Path: Bridging Initial Prejudice and Trainability

2025-05-17 · Alberto Bassi, Carlo Albert, Aurelien Lucchi, Marco Baity-Jesi 외

Understanding the statistical properties of deep neural networks (DNNs) at initialization is crucial for elucidating both their trainability and the intrinsic architectural biases they encode prior to data exposure. Mean…