paper-with-me

홈 › Papers

xRFM: Accurate, scalable, and interpretable feature learning models for tabular data

2025-08-12 · Daniel Beaglehole, David Holzmüller, Adityanarayanan Radhakrishnan, Mikhail Belkin arxiv

Inference from tabular data, collections of continuous and categorical variables organized into matrices, is a foundation for modern technology and science. Yet, in contrast to the explosive changes in the rest of AI, the best practice for these predictive tasks has been relatively unchanged and is still primarily based on variations of Gradient Boosted Decision Trees (GBDTs). Very recently, there has been renewed interest in developing state-of-the-art methods for tabular data based on recent developments in neural networks and feature learning methods. In this work, we introduce xRFM, an algorithm that combines feature learning kernel machines with a tree structure to both adapt to the local structure of the data and scale to essentially unlimited amounts of training data. We show that compared to $31$ other methods, including recently introduced tabular foundation models (TabPFNv2) and GBDTs, xRFM achieves best performance across $100$ regression datasets and is competitive to the best methods across $200$ classification datasets outperforming GBDTs. Additionally, xRFM provides interpretability natively through the Average Gradient Outer Product.

📄 PDF Abstract BibTeX arXiv:2508.10053

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Table2Image: Interpretable Tabular Data Classification with Realistic Image Transformations

2024-12-09 · Seungeun Lee, Il-Youp Kwak, Kihwan Lee, Subin Bae 외

Recent advancements in deep learning for tabular data have shown promise, but challenges remain in achieving interpretable and lightweight models. This paper introduces Table2Image, a novel framework that transforms tabu…

Deep Learning

Barttender: An approachable & interpretable way to compare medical imaging and non-imaging data

2024-11-19 · Ayush Singla, Shakson Isaac, Chirag J. Patel

Imaging-based deep learning has transformed healthcare research, yet its clinical adoption remains limited due to challenges in comparing imaging models with traditional non-imaging and tabular data. To bridge this gap, …

Deep LearningDisease Prediction

Selecting Feature Interactions for Generalized Additive Models by Distilling Foundation Models

2026-04-14 · Jingyun Jia, Chandan Singh, Rich Caruana, Ben Lengerich arxiv

Identifying meaningful feature interactions is a central challenge in building accurate and interpretable models for tabular data. Generalized additive models (GAMs) have shown great success at modeling tabular data, but…

Representation Learning

A Note on Statistically Accurate Tabular Data Generation Using Large Language Models

2025-05-05 · Andrey Sidorenko

Large language models (LLMs) have shown promise in synthetic tabular data generation, yet existing methods struggle to preserve complex feature dependencies, particularly among categorical variables. This work introduces…

Tabular Data Generation

Interpretable Medical Diagnostics with Structured Data Extraction by Large Language Models

2023-06-08 · Aleksa Bisercic, Mladen Nikolic, Mihaela van der Schaar, Boris Delibasic 외

Tabular data is often hidden in text, particularly in medical diagnostic reports. Traditional machine learning (ML) models designed to work with tabular data, cannot effectively process information in such form. On the o…

Diagnostictext-classificationText Classification