paper-with-me

홈 › Papers

Deep Symmetric Autoencoders from the Eckart-Young-Schmidt Perspective

2025-06-13 · Simone Brivio, Nicola Rares Franco

Deep autoencoders have become a fundamental tool in various machine learning applications, ranging from dimensionality reduction and reduced order modeling of partial differential equations to anomaly detection and neural machine translation. Despite their empirical success, a solid theoretical foundation for their expressiveness remains elusive, particularly when compared to classical projection-based techniques. In this work, we aim to take a step forward in this direction by presenting a comprehensive analysis of what we refer to as symmetric autoencoders, a broad class of deep learning architectures ubiquitous in the literature. Specifically, we introduce a formal distinction between different classes of symmetric architectures, analyzing their strengths and limitations from a mathematical perspective. For instance, we show that the reconstruction error of symmetric autoencoders with orthonormality constraints can be understood by leveraging the well-renowned Eckart-Young-Schmidt (EYS) theorem. As a byproduct of our analysis, we end up developing the EYS initialization strategy for symmetric autoencoders, which is based on an iterated application of the Singular Value Decomposition (SVD). To validate our findings, we conduct a series of numerical experiments where we benchmark our proposal against conventional deep autoencoders, discussing the importance of model design and initialization.

📄 PDF Abstract BibTeX arXiv:2506.11641

Code (1)

briviosimone/sae_eys 공식 구현 pytorch

Tasks

Anomaly DetectionDimensionality ReductionMachine Translation

Similar Papers 제목 키워드 기반

Geometry and Optimization of Shallow Polynomial Networks

2025-01-10 · Yossi Arjevani, Joan Bruna, Joe Kileel, Elzbieta Polak 외

We study shallow neural networks with polynomial activations. The function space for these models can be identified with a set of symmetric tensors with bounded rank. We describe general features of these networks, focus…

Spectral Perturbation Bounds for Low-Rank Approximation with Applications to Privacy

2025-10-29 · Phuc Tran, Nisheeth K. Vishnoi, Van H. Vu arxiv

A central challenge in machine learning is to understand how noise or measurement errors affect low-rank approximations, particularly in the spectral norm. This question is especially important in differentially private …

The radius of statistical efficiency

2024-05-15 · Joshua Cutler, Mateo Díaz, Dmitriy Drusvyatskiy

Classical results in asymptotic statistics show that the Fisher information matrix controls the difficulty of estimating a statistical model from observed data. In this work, we introduce a companion measure of robustnes…

Matrix CompletionRetrieval

Neural Network Layer Matrix Decomposition reveals Latent Manifold Encoding and Memory Capacity

2023-09-12 · Ng Shyh-Chang, A-Li Luo, Bo Qiu

We prove the converse of the universal approximation theorem, i.e. a neural network (NN) encoding theorem which shows that for every stably converged NN of continuous activation functions, its weight matrix actually enco…

Compressing Neural Networks: Towards Determining the Optimal Layer-wise Decomposition

2021-07-23 · NeurIPS 2021 12 · Lucas Liebenwein, Alaa Maalouf, Oren Gal, Dan Feldman 외

We present a novel global compression framework for deep neural networks that automatically analyzes each layer to identify the optimal per-layer compression ratio, while simultaneously achieving the desired overall comp…

Low-rank compression