paper-with-me

홈 › Papers

Dual-sPLS: a family of Dual Sparse Partial Least Squares regressions for feature selection and prediction with tunable sparsity; evaluation on simulated and near-infrared (NIR) data

2023-01-17 · Louna Alsouki, Laurent Duval, Clément Marteau, Rami El Haddad, François Wahl

Relating a set of variables X to a response y is crucial in chemometrics. A quantitative prediction objective can be enriched by qualitative data interpretation, for instance by locating the most influential features. When high-dimensional problems arise, dimension reduction techniques can be used. Most notable are projections (e.g. Partial Least Squares or PLS ) or variable selections (e.g. lasso). Sparse partial least squares combine both strategies, by blending variable selection into PLS. The variant presented in this paper, Dual-sPLS, generalizes the classical PLS1 algorithm. It provides balance between accurate prediction and efficient interpretation. It is based on penalizations inspired by classical regression methods (lasso, group lasso, least squares, ridge) and uses the dual norm notion. The resulting sparsity is enforced by an intuitive shrinking ratio parameter. Dual-sPLS favorably compares to similar regression methods, on simulated and real chemical data. Code is provided as an open-source package in R: \url{https://CRAN.R-project.org/package=dual.spls}.

📄 PDF Abstract BibTeX arXiv:2301.07206

Code (1)

https://cran.r-project.org/web/packages/dual.spls 공식 구현

Tasks

Dimensionality Reductionfeature selectionregressionVariable Selection

Similar Papers 제목 키워드 기반

Weighted Sparse Partial Least Squares for Joint Sample and Feature Selection

2023-08-13 · Wenwen Min, Taosheng Xu, Chris Ding

Sparse Partial Least Squares (sPLS) is a common dimensionality reduction technique for data fusion, which projects data samples from two views by seeking linear combinations with a small number of variables with the maxi…

Dimensionality Reductionfeature selection

Supervised Learning for Multi-Block Incomplete Data

2019-01-14 · Hadrien Lorenzo, Jérôme Saracco, Rodolphe Thiébaut

In the supervised high dimensional settings with a large number of variables and a low number of individuals, one objective is to select the relevant variables and thus to reduce the dimension. That subspace selection is…

ImputationVariable Selection

Structured and sparse partial least squares coherence for multivariate cortico-muscular analysis

2025-03-25 · Jingyao Sun, Qilu Zhang, Di Ma, Tianyu Jia 외

Multivariate cortico-muscular analysis has recently emerged as a promising approach for evaluating the corticospinal neural pathway. However, current multivariate approaches encounter challenges such as high dimensionali…

Enhancing software product lines with machine learning components

2025-10-31 · Luz-Viviana Cobaleda, Julián Carvajal, Paola Vallejo, Andrés López 외 arxiv

Modern software systems increasingly integrate machine learning (ML) due to its advancements and ability to enhance data-driven decision-making. However, this integration introduces significant challenges for software en…

Hybrid topic modelling for computational close reading: Mapping narrative themes in Pushkin's Evgenij Onegin

2026-03-20 · Angelo Maria Sabatini arxiv

This study presents a hybrid topic modelling framework for computational literary analysis that integrates Latent Dirichlet Allocation (LDA) with sparse Partial Least Squares Discriminant Analysis (sPLS-DA) to model them…