paper-with-me

Papers

Model as Loss: A Self-Consistent Training Paradigm

2025-05-27 · Saisamarth Rajesh Phaye, Milos Cernak, Andrew Harper

Conventional methods for speech enhancement rely on handcrafted loss functions (e.g., time or frequency domain losses) or deep feature losses (e.g., using WavLM or wav2vec), which often fail to capture subtle signal properties essential for optimal performance. To address this, we propose Model as Loss, a novel training paradigm that utilizes the encoder from the same model as a loss function to guide the training. The Model as Loss paradigm leverages the encoder's task-specific feature space, optimizing the decoder to produce output consistent with perceptual and task-relevant characteristics of the clean signal. By using the encoder's learned features as a loss function, this framework enforces self-consistency between the clean reference speech and the enhanced model output. Our approach outperforms pre-trained deep feature losses on standard speech enhancement benchmarks, offering better perceptual quality and robust generalization to both in-domain and out-of-domain datasets.

📄 PDF Abstract BibTeX arXiv:2505.21156

Code (0)

등록된 구현이 없습니다.

Tasks

DecoderSpeech Enhancement

Similar Papers 제목 키워드 기반

Rethinking Self-Supervised Learning: Small is Beautiful

2021-03-25 · Yun-Hao Cao, Jianxin Wu

Self-supervised learning (SSL), in particular contrastive learning, has made great progress in recent years. However, a common theme in these methods is that they inherit the learning paradigm from the supervised deep le…

Contrastive LearningSelf-Supervised Learning

Pseudo-Stereo Inputs: A Solution to the Occlusion Challenge in Self-Supervised Stereo Matching

2024-10-03 · Ruizhi Yang, Xingqiang Li, Jiajun Bai, Jinsong Du

Self-supervised stereo matching holds great promise for application and research due to its independence from expensive labeled data. However, direct self-supervised stereo matching paradigms based on photometric loss fu…

Stereo Matching

Task Agnostic Representation Consolidation: a Self-supervised based Continual Learning Approach

2022-07-13 · Prashant Bhat, Bahram Zonooz, Elahe Arani

Continual learning (CL) over non-stationary data streams remains one of the long-standing challenges in deep neural networks (DNNs) as they are prone to catastrophic forgetting. CL models can benefit from self-supervised…

Continual Learning

How Do Electrocardiogram Models Scale?

2026-05-17 · Jiawei Li, Fabio Bonassi, Ming Jin, Stefan Gustafsson 외 arxiv

While scaling laws have established a fundamental framework for foundation models in natural language processing, their applicability to electrocardiogram (ECG) models remains poorly characterized. Indeed, recent studies…

Self-Supervised Learning

SSM-Net: feature learning for Music Structure Analysis using a Self-Similarity-Matrix based loss

2022-11-15 · Geoffroy Peeters, Florian Angulo

In this paper, we propose a new paradigm to learn audio features for Music Structure Analysis (MSA). We train a deep encoder to learn features such that the Self-Similarity-Matrix (SSM) resulting from those approximates …