paper-with-me

Papers

Metadata Normalization

2021-04-19 · CVPR 2021 1 · Mandy Lu, Qingyu Zhao, Jiequan Zhang, Kilian M. Pohl, Li Fei-Fei, Juan Carlos Niebles, Ehsan Adeli

Batch Normalization (BN) and its variants have delivered tremendous success in combating the covariate shift induced by the training step of deep learning methods. While these techniques normalize feature distributions by standardizing with batch statistics, they do not correct the influence on features from extraneous variables or multiple distributions. Such extra variables, referred to as metadata here, may create bias or confounding effects (e.g., race when classifying gender from face images). We introduce the Metadata Normalization (MDN) layer, a new batch-level operation which can be used end-to-end within the training framework, to correct the influence of metadata on feature distributions. MDN adopts a regression analysis technique traditionally used for preprocessing to remove (regress out) the metadata effects on model features during training. We utilize a metric based on distance correlation to quantify the distribution bias from the metadata and demonstrate that our method successfully removes metadata effects on four diverse settings: one synthetic, one 2D image, one video, and one 3D medical image dataset.

📄 PDF Abstract BibTeX arXiv:2104.09052

Code (1)

mlu355/MetadataNorm 공식 구현 pytorch

Similar Papers 제목 키워드 기반

Sensor-Type Classification in Buildings

2015-09-01 · Dezhi Hong, Jorge Ortiz, Arka Bhattacharya, Kamin Whitehouse

Many sensors/meters are deployed in commercial buildings to monitor and optimize their performance. However, because sensor metadata is inconsistent across buildings, software-based solutions are tightly coupled to the s…

ClassificationEnsemble LearningGeneral ClassificationVocal Bursts Type Prediction

Structure-Aware RAG: Structured Retrieval Augmented Generation from Noisy Data for Conversational Agents

2026-05-23 · Kaiqiao Han, LuAn Tang, Renliang Sun, Peng Yuan 외 arxiv

Large Language Models (LLMs) have been widely adopted in conversational applications. However, their reliance on parametric knowledge limits reliability in real-world scenarios that require dynamic or domain-specific inf…

A Penalty Approach for Normalizing Feature Distributions to Build Confounder-Free Models

2022-07-11 · Anthony Vento, Qingyu Zhao, Robert Paul, Kilian M. Pohl 외

Translating machine learning algorithms into clinical applications requires addressing challenges related to interpretability, such as accounting for the effect of confounding variables (or metadata). Confounding variabl…

Enhancing Omics Cohort Discovery for Research on Neurodegeneration through Ontology-Augmented Embedding Models

2025-06-16 · José A. Pardo, Alicia Gómez-Pascual, José T. Palma, Juan A. Botía

The growing volume of omics and clinical data generated for neurodegenerative diseases (NDs) requires new approaches for their curation so they can be ready-to-use in bioinformatics. NeuroEmbed is an approach for the eng…

Question Answering

CAMEO: Collection of Multilingual Emotional Speech Corpora

2025-05-16 · Iwona Christop, Maciej Czajka

This paper presents CAMEO -- a curated collection of multilingual emotional speech datasets designed to facilitate research in emotion recognition and other speech-related tasks. The main objectives were to ensure easy a…

Emotion RecognitionSpeech Emotion Recognition