paper-with-me

Speech Extraction

1개 벤치마크 · 논문 55편 · 이 태스크의 논문 보기 →

Benchmarks

WSJ0-2mix-extr

결과 1개

Most implemented

Papers

Breaking Shortcut Learning for Cross-Trial EEG-Guided Target Speech Extraction via Two-Stage Training

2026-06-23 · Wonchul Shin, Inyong Choi, Kyogu Lee arxiv

Recent end-to-end models for EEG-guided target speech extraction report impressive results, underscoring potential for neuro-steered hearing technologies. However, our analysis reveals that high within-trial performance …

Speech Extraction

IsoNet: Spatially-aware audio-visual target speech extraction in complex acoustic environments

2026-05-14 · Dinanath Padhya, Sajen Maharjan, Binita Adhikari, Ishwor Raj Pokharel arxiv

Target speech extraction remains difficult for compact devices because monaural neural models lack spatial evidence and classical beamformers lose resolving power when the microphone aperture is only a few centimetres. W…

Speech Extraction

VorTEX: Various overlap ratio for Target speech EXtraction

2026-03-16 · Ro-hoon Oh, Jihwan Seol, Bugeun Kim arxiv

Target speech extraction (TSE) aims to recover a target speaker's voice from a mixture. While recent text-prompted approaches have shown promise, most approaches assume fully overlapped mixtures, limiting insight into be…

Speech Extraction

Lightweight Wasserstein Audio-Visual Model for Unified Speech Enhancement and Separation

2025-12-07 · Jisoo Park, Seonghak Lee, Guisik Kim, Taewoo Kim 외 arxiv

Speech Enhancement (SE) and Speech Separation (SS) have traditionally been treated as distinct tasks in speech processing. However, real-world audio often involves both background noise and overlapping speakers, motivati…

Speech EnhancementSpeech ExtractionSpeech Separation

ELEGANCE: Efficient LLM Guidance for Audio-Visual Target Speech Extraction

2025-11-09 · Wenxuan Wu, Shuai Wang, Xixin Wu, Helen Meng 외 arxiv

Audio-visual target speaker extraction (AV-TSE) models primarily rely on visual cues from the target speaker. However, humans also leverage linguistic knowledge, such as syntactic constraints, next word prediction, and p…

Speech Extraction

Neural Speech Extraction with Human Feedback

2025-08-05 · Malek Itani, Ashton Graves, Sefik Emre Eskimez, Shyamnath Gollakota arxiv

We present the first neural target speech extraction (TSE) system that uses human feedback for iterative refinement. Our approach allows users to mark specific segments of the TSE output, generating an edit mask. The ref…

Speech Extraction

전체 55편 보기 →