paper-with-me

Papers

Learning Discrete Concepts in Latent Hierarchical Models

2024-06-01 · Lingjing Kong, Guangyi Chen, Biwei Huang, Eric P. Xing, Yuejie Chi, Kun Zhang

Learning concepts from natural high-dimensional data (e.g., images) holds potential in building human-aligned and interpretable machine learning models. Despite its encouraging prospect, formalization and theoretical insights into this crucial task are still lacking. In this work, we formalize concepts as discrete latent causal variables that are related via a hierarchical causal model that encodes different abstraction levels of concepts embedded in high-dimensional data (e.g., a dog breed and its eye shapes in natural images). We formulate conditions to facilitate the identification of the proposed causal model, which reveals when learning such concepts from unsupervised data is possible. Our conditions permit complex causal hierarchical structures beyond latent trees and multi-level directed acyclic graphs in prior work and can handle high-dimensional, continuous observed variables, which is well-suited for unstructured data modalities such as images. We substantiate our theoretical claims with synthetic data experiments. Further, we discuss our theory's implications for understanding the underlying mechanisms of latent diffusion models and provide corresponding empirical evidence for our theoretical insights.

📄 PDF Abstract BibTeX arXiv:2406.00519

Code (0)

등록된 구현이 없습니다.

Tasks

Interpretable Machine Learning

Methods 이 논문이 사용한 방법론

Diffusion Diffusion models generate samples by gradually removing noise from a signal, and their training objective can be expressed as a reweighted variational lower-bound…

Similar Papers 제목 키워드 기반

Vector Quantized Latent Concepts: A Scalable Alternative to Clustering-Based Concept Discovery

2026-02-02 · Xuemin Yu, Ankur Garg, Samira Ebrahimi Kahou, Hassan Sajjad arxiv

Large language models (LLMs) encode rich semantic information in their hidden states, yet it remains difficult to understand what information these internal representations capture. Latent concepts extracted from hidden …

Discrete Causal Representations from Heterogeneous Domains: A Bayesian Approach with Social Survey Applications

2026-06-04 · Ankur Garg, Michael Stettler, Aaron Schein, Julius von Kügelgen arxiv

Causal representation learning aims to infer the high-level latent causal concepts that give rise to observed low-level measurements. This is particularly relevant for heterogeneous data from different environments or do…

Representation Learning

Relaxed-Responsibility Hierarchical Discrete VAEs

2020-07-14 · Matthew Willetts, Xenia Miscouridou, Stephen Roberts, Chris Holmes

Successfully training Variational Autoencoders (VAEs) with a hierarchy of discrete latent variables remains an area of active research. Vector-Quantised VAEs are a powerful approach to discrete VAEs, but naive hierarchic…

Hyperbolic Residual Quantization: Discrete Representations for Data with Latent Hierarchies

2025-05-18 · Piotr Piękos, Subhradeep Kayal, Alexandros Karatzoglou

Hierarchical data arise in countless domains, from biological taxonomies and organizational charts to legal codes and knowledge graphs. Residual Quantization (RQ) is widely used to generate discrete, multitoken represent…

Inductive BiasKnowledge GraphsQuantizationRepresentation Learning

Improving Latent Reasoning in LLMs via Soft Concept Mixing

2025-11-21 · Kang Wang, Xiangyu Duan, Tianyi Du arxiv

Unlike human reasoning in abstract conceptual spaces, large language models (LLMs) typically reason by generating discrete tokens, which potentially limit their expressive power. The recent work Soft Thinking has shown t…

Reinforcement Learning