paper-with-me

Papers

Deep Sufficient Representation Learning via Mutual Information

2022-07-21 · Siming Zheng, Yuanyuan Lin, Jian Huang

We propose a mutual information-based sufficient representation learning (MSRL) approach, which uses the variational formulation of the mutual information and leverages the approximation power of deep neural networks. MSRL learns a sufficient representation with the maximum mutual information with the response and a user-selected distribution. It can easily handle multi-dimensional continuous or categorical response variables. MSRL is shown to be consistent in the sense that the conditional probability density function of the response variable given the learned representation converges to the conditional probability density function of the response variable given the predictor. Non-asymptotic error bounds for MSRL are also established under suitable conditions. To establish the error bounds, we derive a generalized Dudley's inequality for an order-two U-process indexed by deep neural networks, which may be of independent interest. We discuss how to determine the intrinsic dimension of the underlying data distribution. Moreover, we evaluate the performance of MSRL via extensive numerical experiments and real data analysis and demonstrate that MSRL outperforms some existing nonlinear sufficient dimension reduction methods.

📄 PDF Abstract BibTeX arXiv:2207.10772

Code (0)

등록된 구현이 없습니다.

Tasks

Dimensionality ReductionRepresentation Learning

Similar Papers 제목 키워드 기반

Which Mutual-Information Representation Learning Objectives are Sufficient for Control?

2021-06-14 · NeurIPS 2021 12 · Kate Rakelly, Abhishek Gupta, Carlos Florensa, Sergey Levine

Mutual information maximization provides an appealing formalism for learning representations of data. In the context of reinforcement learning (RL), such representations can accelerate learning by discarding irrelevant a…

Reinforcement Learning (RL)Representation Learning

Trimming the Independent Fat: Sufficient Statistics, Mutual Information, and Predictability from Effective Channel States

2017-02-07 · Ryan G. James, John R. Mahoney, James P. Crutchfield

One of the most fundamental questions one can ask about a pair of random variables X and Y is the value of their mutual information. Unfortunately, this task is often stymied by the extremely large dimension of the varia…

Matching Text with Deep Mutual Information Estimation

2020-03-09 · Xixi Zhou, Chengxi Li, Jiajun Bu, Chengwei Yao 외

Text matching is a core natural language processing research problem. How to retain sufficient information on both content and structure information is one important challenge. In this paper, we present a neural approach…

Answer SelectionMutual Information EstimationNatural Language InferenceParaphrase Identification+1

MVEB: Self-Supervised Learning with Multi-View Entropy Bottleneck

2024-03-28 · Liangjian Wen, Xiasi Wang, Jianzhuang Liu, Zenglin Xu

Self-supervised learning aims to learn representation that can be effectively generalized to downstream tasks. Many self-supervised approaches regard two views of an image as both the input and the self-supervised signal…

Linear evaluationSelf-Supervised Learning

InfoPrompt: Information-Theoretic Soft Prompt Tuning for Natural Language Understanding

2023-06-08 · NeurIPS 2023 11

Soft prompt tuning achieves superior performances across a wide range of few-shot tasks. However, the performances of prompt tuning can be highly sensitive to the initialization of the prompts. We also empirically observ…

Language ModelingLanguage ModellingNatural Language Understanding