paper-with-me

홈 › Papers

Twoblock clustering trees with coskewness-based dimension reduction: recovering piecewise multivariate linear regimes

2026-07-22 · Sven Serneels arxiv

The twoblock clustering tree (\tbtree) is introduced as a highly interpretable regression tree for multivariate responses. Twoblock trees are deterministic decision trees that have local multivariate linear models as their leaves and use dense or sparse twoblock dimension reduction as local leaf models and in the impurity. The resulting models are both computationally efficient and can be highly interpretable. Beyond proposing the decision tree estimator itself, this paper also introduces an estimator for the twoblock dimension reduced space based on maximizing coskewness, which facilitates identification of non-normal clusters in the data. The tree inherently produces a set of local linear models and is therefore apt to recover peicewise linear regimes, which is illustrated in a simulation. However, two real world data examples illustrate that twoblock trees are also capable of modeling more complexly nonlinear dependencies and can perform on par with black box modeling techniques, such as random forests. At each point, both the twoblock models that generate the splits, as well as the ones in the leaves, can be inspected and interpreted.

📄 PDF Abstract BibTeX arXiv:2607.20760

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Multi Loci Phylogenetic Analysis with Gene Tree Clustering

2015-12-31

Summary: Both theory and empirical evidence indicate that phylogenies (trees) of different genes (loci) do not display precisely matched topologies. This phylogenetic incongruence is attributed to the reticulated evoluti…

ClusteringDimensionality Reduction

Scalable Differentially Private Clustering via Hierarchically Separated Trees

2022-06-17 · Vincent Cohen-Addad, Alessandro Epasto, Silvio Lattanzi, Vahab Mirrokni 외

We study the private $k$-median and $k$-means clustering problem in $d$ dimensional Euclidean space. By leveraging tree embeddings, we give an efficient and easy to implement algorithm, that is empirically competitive wi…

ClusteringDimensionality ReductionDistributed Computing

Assessing the impact of dimensionality reduction on clustering performance -- a systematic study

2026-04-23 · Ousmane Assani-Amate, Mohammadreza Bakhtyari, Émilie Roy, Vladimir Makarenkov arxiv

Dimensionality reduction is a critical preprocessing step for clustering high-dimensional data, yet comprehensive evaluation of its impact across diverse methods and data types remains limited. In this study, we systemat…

Dimensionality Reduction

Self-supervising Action Recognition by Statistical Moment and Subspace Descriptors

2020-01-14 · Lei Wang, Piotr Koniusz

In this paper, we build on a concept of self-supervision by taking RGB frames as input to learn to predict both action concepts and auxiliary descriptors e.g., object descriptors. So-called hallucination streams are trai…

Action ClassificationAction RecognitionEgocentric Activity RecognitionHallucination+2

Scalable Bottom-up Subspace Clustering using FP-Trees for High Dimensional Data

2018-11-07 · Minh Tuan Doan, Jianzhong Qi, Sutharshan Rajasegarar, Christopher Leckie

Subspace clustering aims to find groups of similar objects (clusters) that exist in lower dimensional subspaces from a high dimensional dataset. It has a wide range of applications, such as analysing high dimensional sen…

Clustering