paper-with-me

Papers

SMART Fine-tuning Factor Augmented Neural Lasso

2026-04-14 · Jinhang Chai, Jianqing Fan, Cheng Gao, Qishuo Yin arxiv

Fine-tuning is a widely used strategy for adapting pre-trained models to new tasks, yet its methodology and theoretical properties in high-dimensional nonparametric settings with variable selection have not yet been developed. We propose a source-model-augmented residual tuning (SMART) framework, which incorporates the pre-trained source model as an augmented feature into the target learner and estimates only the residual target-specific component. The approach is widely applicable, from parametric and sparse models to neural networks and blackbox machine learning models. We focus on the development of fine-tuning factor-augmented neural Lasso, resulting in SMART-FAN-Lasso. This transfer-learning framework for high-dimensional nonparametric regression with variable selection simultaneously handles covariate and posterior shifts. We use a low-rank factor structure to manage high-dimensional dependent covariates and a residual tuning decomposition in which the target function is expressed as a function of source model and other target-specific variables, thereby reducing the effective complexity of the target task. We derive minimax-optimal excess risk bounds, characterizing the precise conditions, in terms of relative sample sizes and function complexities, under which fine-tuning yields statistical acceleration over single-task learning. Extensive numerical experiments across diverse covariate- and posterior-shift scenarios demonstrate that SMART-FAN-Lasso consistently outperforms standard baselines and achieves near-oracle performance even under severe target sample size constraints, empirically validating the derived rates.

📄 PDF Abstract BibTeX arXiv:2604.12288

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Factor-Augmented Machine Learning Panel Regressions

2026-07-07 · Andrii Babii, Luca Barbaglia, Eric Ghysels, Jonas Striaukas arxiv

This paper develops the asymptotic theory for high-dimensional panel data regressions in settings with cross-sectionally dependent errors driven by common shocks. We consider a factor-augmented sparse-group LASSO estimat…

Addressing Data Distribution Shifts in Online Machine Learning Powered Smart City Applications Using Augmented Test-Time Adaptation

2022-11-02 · Shawqi Al-Maliki, Faissal El Bouanani, Mohamed Abdallah, Junaid Qadir 외

Data distribution shift is a common problem in machine learning-powered smart city applications where the test data differs from the training data. Augmenting smart city applications with online machine learning models c…

Test-time Adaptation

Domain-Specific Retrieval-Augmented Generation Using Vector Stores, Knowledge Graphs, and Tensor Factorization

2024-10-03 · Ryan C. Barron, Ves Grantcharov, Selma Wanna, Maksim E. Eren 외

Large Language Models (LLMs) are pre-trained on large-scale corpora and excel in numerous general natural language processing (NLP) tasks, such as question answering (QA). Despite their advanced language capabilities, wh…

Anomaly DetectionAttributeKnowledge GraphsMalware Analysis+5

Risk-consistency of cross-validation with lasso-type procedures

2013-08-04 · Darren Homrighausen, Daniel J. McDonald

The lasso and related sparsity inducing algorithms have been the target of substantial theoretical and applied research. Correspondingly, many results are known about their behavior for a fixed or optimally chosen tuning…

Model SelectionVocal Bursts Type Prediction

LLM-Lasso: A Robust Framework for Domain-Informed Feature Selection and Regularization

2025-02-15 · Erica Zhang, Ryunosuke Goto, Naomi Sagan, Jurik Mutter 외

We introduce LLM-Lasso, a novel framework that leverages large language models (LLMs) to guide feature selection in Lasso $\ell_1$ regression. Unlike traditional methods that rely solely on numerical data, LLM-Lasso inco…

feature selectionRAGRetrieval-augmented Generation