paper-with-me

Papers

Random Bias Initialization Improves Quantized Training

2019-09-30 · Xinlin Li, Vahid Partovi Nia

Binary neural networks improve computationally efficiency of deep models with a large margin. However, there is still a performance gap between a successful full-precision training and binary training. We bring some insights about why this accuracy drop exists and call for a better understanding of binary network geometry. We start with analyzing full-precision neural networks with ReLU activation and compare it with its binarized version. This comparison suggests to initialize networks with random bias, a counter-intuitive remedy.

📄 PDF Abstract BibTeX arXiv:1909.13446

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

ReLU How Do I Communicate to Expedia? How Do I Communicate to Expedia? – Call ☎️ +1-(888) 829 (0881) or +1-805-330-4056 or +1-805-330-4056 for Live Support & Special Travel…

Similar Papers 제목 키워드 기반

Starting Positions Matter: A Study on Better Weight Initialization for Neural Network Quantization

2025-06-12 · Stone Yun, Alexander Wong

Deep neural network (DNN) quantization for fast, efficient inference has been an important tool in limiting the cost of machine learning (ML) model inference. Quantization-specific model development techniques such as re…

Quantization

Target noise: A pre-training based neural network initialization for efficient high resolution learning

2026-02-06 · Shaowen Wang, Tariq Alkhalifah arxiv

Weight initialization plays a crucial role in the optimization behavior and convergence efficiency of neural networks. Most existing initialization methods, such as Xavier and Kaiming initializations, rely on random samp…

GHN-QAT: Training Graph Hypernetworks to Predict Quantization-Robust Parameters of Unseen Limited Precision Neural Networks

2023-09-24 · Stone Yun, Alexander Wong

Graph Hypernetworks (GHN) can predict the parameters of varying unseen CNN architectures with surprisingly good accuracy at a fraction of the cost of iterative optimization. Following these successes, preliminary researc…

Quantization

Transformers Are Born Biased: Structural Inductive Biases at Random Initialization and Their Practical Consequences

2026-02-05 · Siquan Li, Yao Tong, Haonan Wang, Tianyang Hu arxiv

Transformers underpin modern large language models (LLMs) and are commonly assumed to be behaviorally unstructured at random initialization, with all meaningful preferences emerging only through large-scale training. We …

Amenable Sparse Network Investigator

2022-02-18 · Saeed Damadi, Erfan Nouri, Hamed Pirsiavash

We present "Amenable Sparse Network Investigator" (ASNI) algorithm that utilizes a novel pruning strategy based on a sigmoid function that induces sparsity level globally over the course of one single round of training. …

Quantization