Study Features via Exploring Distribution Structure
In this paper, we present a novel framework for data redundancy measurement based on probabilistic modeling of datasets, and a new criterion for redundancy detection that is resilient to noise. We also develop new methods for data redundancy reduction using both deterministic and stochastic optimization techniques. Our framework is flexible and can handle different types of features, and our experiments on benchmark datasets demonstrate the effectiveness of our methods. We provide a new perspective on feature selection, and propose effective and robust approaches for both supervised and unsupervised learning problems.
Code (0)
등록된 구현이 없습니다.
Tasks
feature selectionStochastic OptimizationSimilar Papers 제목 키워드 기반
Feasible Joint Posterior Beliefs
We study the set of possible joint posterior belief distributions of a group of agents who share a common prior regarding a binary state, and who observe some information structure. For two agents we introduce a quantita…
Exploring Optimal Substructure for Out-of-distribution Generalization via Feature-targeted Model Pruning
Recent studies show that even highly biased dense networks contain an unbiased substructure that can achieve better out-of-distribution (OOD) generalization than the original model. Existing works usually search the inva…
Out-of-Distribution GeneralizationExploring the Representational Power of Graph Autoencoder
While representation learning has yielded a great success on many graph learning tasks, there is little understanding behind the structures that are being captured by these embeddings. For example, we wonder if the topol…
ClusteringGraph EmbeddingGraph LearningRepresentation LearningExploring the Limitations of kNN Noisy Feature Detection and Recovery for Self-Driving Labs
Self-driving laboratories (SDLs) have shown promise to accelerate materials discovery by integrating machine learning with automated experimental platforms. However, errors in the capture of input parameters may corrupt …
SiT: Exploring Flow and Diffusion-based Generative Models with Scalable Interpolant Transformers
We present Scalable Interpolant Transformers (SiT), a family of generative models built on the backbone of Diffusion Transformers (DiT). The interpolant framework, which allows for connecting two distributions in a more …
Image Generation