paper-with-me

홈 › Papers

A Probabilistic Model for Data Redundancy in the Feature Domain

2023-09-24 · Ghurumuruhan Ganesan

In this paper, we use a probabilistic model to estimate the number of uncorrelated features in a large dataset. Our model allows for both pairwise feature correlation (collinearity) and interdependency of multiple features (multicollinearity) and we use the probabilistic method to obtain upper and lower bounds of the same order, for the size of a feature set that exhibits low collinearity and low multicollinearity. We also prove an auxiliary result regarding mutually good constrained sets that is of independent interest.

📄 PDF Abstract BibTeX arXiv:2309.13657

Code (0)

등록된 구현이 없습니다.

Tasks

Feature Correlation

Similar Papers 제목 키워드 기반

Study Features via Exploring Distribution Structure

2024-01-15 · Chunxu Cao, Qiang Zhang

In this paper, we present a novel framework for data redundancy measurement based on probabilistic modeling of datasets, and a new criterion for redundancy detection that is resilient to noise. We also develop new method…

feature selectionStochastic Optimization

A New Hierarchical Redundancy Eliminated Tree Augmented Naive Bayes Classifier for Coping with Gene Ontology-based Features

2016-07-06 · Cen Wan, Alex A. Freitas

The Tree Augmented Naive Bayes classifier is a type of probabilistic graphical model that can represent some feature dependencies. In this work, we propose a Hierarchical Redundancy Eliminated Tree Augmented Naive Bayes …

Improving Unsupervised Domain Adaptation by Reducing Bi-level Feature Redundancy

2020-12-28 · Mengzhu Wang, Xiang Zhang, Long Lan, Wei Wang 외

Reducing feature redundancy has shown beneficial effects for improving the accuracy of deep learning models, thus it is also indispensable for the models of unsupervised domain adaptation (UDA). Nevertheless, most recent…

Domain AdaptationUnsupervised Domain Adaptation

Training and Inference on Any-Order Autoregressive Models the Right Way

2022-05-26 · Andy Shih, Dorsa Sadigh, Stefano Ermon

Conditional inference on arbitrary subsets of variables is a core problem in probabilistic inference with important applications such as masked language modeling and image inpainting. In recent years, the family of Any-O…

Image InpaintingLanguage ModelingLanguage ModellingMasked Language Modeling

Hierarchical Dependency Constrained Tree Augmented Naive Bayes Classifiers for Hierarchical Feature Spaces

2022-02-08 · Cen Wan, Alex A. Freitas

The Tree Augmented Naive Bayes (TAN) classifier is a type of probabilistic graphical model that constructs a single-parent dependency tree to estimate the distribution of the data. In this work, we propose two novel Hier…