paper-with-me

홈 › Papers

Learning Without Training

2026-02-20 · Ryan O'Dowd arxiv

Machine learning is at the heart of managing the real-world problems associated with massive data. With the success of neural networks on such large-scale problems, more research in machine learning is being conducted now than ever before. This dissertation focuses on three different projects rooted in mathematical theory for machine learning applications. The first project deals with supervised learning and manifold learning. In theory, one of the main problems in supervised learning is that of function approximation: that is, given some data set $\mathcal{D}=\{(x_j,f(x_j))\}_{j=1}^M$, can one build a model $F\approx f$? We introduce a method which aims to remedy several of the theoretical shortcomings of the current paradigm for supervised learning. The second project deals with transfer learning, which is the study of how an approximation process or model learned on one domain can be leveraged to improve the approximation on another domain. We study such liftings of functions when the data is assumed to be known only on a part of the whole domain. We are interested in determining subsets of the target data space on which the lifting can be defined, and how the local smoothness of the function and its lifting are related. The third project is concerned with the classification task in machine learning, particularly in the active learning paradigm. Classification has often been treated as an approximation problem as well, but we propose an alternative approach leveraging techniques originally introduced for signal separation problems. We introduce theory to unify signal separation with classification and a new algorithm which yields competitive accuracy to other recent active learning algorithms while providing results much faster.

📄 PDF Abstract BibTeX arXiv:2602.17985

Code (0)

등록된 구현이 없습니다.

Tasks

Transfer LearningActive Learning

Similar Papers 제목 키워드 기반

Influence Function Based Second-Order Channel Pruning-Evaluating True Loss Changes For Pruning Is Possible Without Retraining

2023-08-13 · Hongrong Cheng, Miao Zhang, Javen Qinfeng Shi

A challenge of channel pruning is designing efficient and effective criteria to select channels to prune. A widely used criterion is minimal performance degeneration. To accurately evaluate the truth performance degenera…

Greedy Output Approximation: Towards Efficient Structured Pruning for LLMs Without Retraining

2024-07-26 · Jianwei Li, Yijun Dong, Qi Lei

To remove redundant components of large language models (LLMs) without incurring significant computational costs, this work focuses on single-shot pruning without a retraining phase. We simplify the pruning process for T…

Practical FP4 Training for Large-Scale MoE Models on Hopper GPUs

2026-03-03 · Wuyue Zhang, Chongdong Huang, Chunbo You, Cheng Gu 외 arxiv

Training large-scale Mixture-of-Experts (MoE) models is bottlenecked by activation memory and expert-parallel communication, yet FP4 training remains impractical on Hopper-class GPUs without native MXFP4 or NVFP4 support…

Training Deep Architectures Without End-to-End Backpropagation: A Survey on the Provably Optimal Methods

2021-01-09 · Shiyu Duan, Jose C. Principe

This tutorial paper surveys provably optimal alternatives to end-to-end backpropagation (E2EBP) -- the de facto standard for training deep architectures. Modular training refers to strictly local training without both th…

ASIF: Coupled Data Turns Unimodal Models to Multimodal Without Training

2022-10-04 · NeurIPS 2023 11 · Antonio Norelli, Marco Fumero, Valentino Maiorca, Luca Moschella 외

CLIP proved that aligning visual and language spaces is key to solving many vision tasks without explicit training, but required to train image and text encoders from scratch on a huge dataset. LiT improved this by only …