paper-with-me

홈 › Papers

Learning Robust Models Using The Principle of Independent Causal Mechanisms

2020-10-14 · Jens Müller, Robert Schmier, Lynton Ardizzone, Carsten Rother, Ullrich Köthe

Standard supervised learning breaks down under data distribution shift. However, the principle of independent causal mechanisms (ICM, Peters et al. (2017)) can turn this weakness into an opportunity: one can take advantage of distribution shift between different environments during training in order to obtain more robust models. We propose a new gradient-based learning framework whose objective function is derived from the ICM principle. We show theoretically and experimentally that neural networks trained in this framework focus on relations remaining invariant across environments and ignore unstable ones. Moreover, we prove that the recovered stable relations correspond to the true causal mechanisms under certain conditions. In both regression and classification, the resulting models generalize well to unseen scenarios where traditionally trained models fail.

📄 PDF Abstract BibTeX arXiv:2010.07167

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Learning Causally Disentangled Representations via the Principle of Independent Causal Mechanisms

2023-06-02 · Aneesh Komanduri, Yongkai Wu, Feng Chen, Xintao Wu

Learning disentangled causal representations is a challenging problem that has gained significant attention recently due to its implications for extracting meaningful information for downstream tasks. In this work, we de…

counterfactualDisentanglement

Causal Direction of Data Collection Matters: Implications of Causal and Anticausal Learning for NLP

2021-10-07 · EMNLP 2021 11 · Zhijing Jin, Julius von Kügelgen, Jingwei Ni, Tejas Vaidhya 외

The principle of independent causal mechanisms (ICM) states that generative processes of real world data consist of independent modules which do not influence or inform each other. While this idea has led to fruitful dev…

Causal InferenceDomain Adaptation

Fail-Closed Alignment for Large Language Models

2026-02-19 · Zachary Coalson, Beth Sohler, Aiden Gabriel, Sanghyun Hong arxiv

We identify a structural weakness in current large language model (LLM) alignment: modern refusal mechanisms are fail-open. While existing approaches encode refusal behaviors across multiple latent features, suppressing …

Independent mechanism analysis, a new concept?

2021-06-09 · NeurIPS 2021 12 · Luigi Gresele, Julius von Kügelgen, Vincent Stimper, Bernhard Schölkopf 외

Independent component analysis provides a principled framework for unsupervised representation learning, with solid theory on the identifiability of the latent code that generated the data, given only observations of mix…

blind source separationRepresentation Learning

Towards Error-Centric Intelligence II: Energy-Structured Causal Models

2025-10-24 · Marcus Thomas arxiv

Contemporary machine learning optimizes for predictive accuracy, yet systems that achieve state of the art performance remain causally opaque: their internal representations provide no principled handle for intervention.…