paper-with-me

홈 › Papers

Principled Approximation Methods for Efficient and Scalable Deep Learning

2025-08-29 · Pedro Savarese arxiv

Recent progress in deep learning has been driven by increasingly larger models. However, their computational and energy demands have grown proportionally, creating significant barriers to their deployment and to a wider adoption of deep learning technologies. This thesis investigates principled approximation methods for improving the efficiency of deep learning systems, with a particular focus on settings that involve discrete constraints and non-differentiability. We study three main approaches toward improved efficiency: architecture design, model compression, and optimization. For model compression, we propose novel approximations for pruning and quantization that frame the underlying discrete problem as continuous and differentiable, enabling gradient-based training of compression schemes alongside the model's parameters. These approximations allow for fine-grained sparsity and precision configurations, leading to highly compact models without significant fine-tuning. In the context of architecture design, we design an algorithm for neural architecture search that leverages parameter sharing across layers to efficiently explore implicitly recurrent architectures. Finally, we study adaptive optimization, revisiting theoretical properties of widely used methods and proposing an adaptive optimizer that allows for quick hyperparameter tuning. Our contributions center on tackling computationally hard problems via scalable and principled approximations. Experimental results on image classification, language modeling, and generative modeling tasks show that the proposed methods provide significant improvements in terms of training and inference efficiency while maintaining, or even improving, the model's performance.

📄 PDF Abstract BibTeX arXiv:2509.00174

Code (0)

등록된 구현이 없습니다.

Tasks

Neural Architecture SearchImage ClassificationModel Compression

Similar Papers 제목 키워드 기반

Random Wavelet Features for Graph Kernel Machines

2026-02-17 · Valentin de Bassompierre, Jean-Charles Delvenne, Laurent Jacques arxiv

Node embeddings map graph vertices into low-dimensional Euclidean spaces while preserving structural information. They are central to tasks such as node classification, link prediction, and signal reconstruction. A key g…

Graph Representation LearningNode ClassificationLink Prediction

Scalable Random Wavelet Features: Efficient Non-Stationary Kernel Approximation with Convergence Guarantees

2026-02-01 · Sawan Kumar, Souvik Chakraborty arxiv

Modeling non-stationary processes, where statistical properties vary across the input domain, is a critical challenge in machine learning; yet most scalable methods rely on a simplifying assumption of stationarity. This …

Gaussian Processes

Structured Variational Inference for Coupled Gaussian Processes

2017-11-03 · Vincent Adam

Sparse variational approximations allow for principled and scalable inference in Gaussian Process (GP) models. In settings where several GPs are part of the generative model, theses GPs are a posteriori coupled. For many…

Gaussian ProcessesVariational Inference

Variational Bayes for Merging Noisy Databases

2014-10-17 · Tamara Broderick, Rebecca C. Steorts

Bayesian entity resolution merges together multiple, noisy databases and returns the minimal collection of unique individuals represented, together with their true, latent record values. Bayesian methods allow flexible g…

Bayesian InferenceEntity Resolution

Reinforcement Learning via AIXI Approximation

2010-07-13 · AAAI 2010 2010 7 · Joel Veness, Kee Siong Ng, Marcus Hutter, David Silver

This paper introduces a principled approach for the design of a scalable general reinforcement learning agent. This approach is based on a direct approximation of AIXI, a Bayesian optimality notion for general reinforcem…

General Reinforcement LearningOpen-Ended Question Answeringreinforcement-learningReinforcement Learning+1