Semi-supervised source localization with deep generative modeling
We propose a semi-supervised localization approach based on deep generative modeling with variational autoencoders (VAEs). Localization in reverberant environments remains a challenge, which machine learning (ML) has shown promise in addressing. Even with large data volumes, the number of labels available for supervised learning in reverberant environments is usually small. We address this issue by performing semi-supervised learning (SSL) with convolutional VAEs. The VAE is trained to generate the phase of relative transfer functions (RTFs), in parallel with a DOA classifier, on both labeled and unlabeled RTF samples. The VAE-SSL approach is compared with SRP-PHAT and fully-supervised CNNs. We find that VAE-SSL can outperform both SRP-PHAT and CNN in label-limited scenarios.
Code (0)
등록된 구현이 없습니다.
Methods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
Semi-supervised source localization in reverberant environments with deep generative modeling
We propose a semi-supervised approach to acoustic source localization in reverberant environments based on deep generative modeling. Localization in reverberant environments remains an open challenge. Even with large dat…
Semi-Supervised Learning with GANs for Device-Free Fingerprinting Indoor Localization
Device-free wireless indoor localization is a key enabling technology for the Internet of Things (IoT). Fingerprint-based indoor localization techniques are a commonly used solution. This paper proposes a semi-supervised…
Generative Adversarial NetworkIndoor LocalizationTwo-stage Denoising Diffusion Model for Source Localization in Graph Inverse Problems
Source localization is the inverse problem of graph information dissemination and has broad practical applications. However, the inherent intricacy and uncertainty in information dissemination pose significant challenges…
DenoisingSemi-Supervised Audio-Visual Video Action Recognition with Audio Source Localization Guided Mixup
Video action recognition is a challenging but important task for understanding and discovering what the video does. However, acquiring annotations for a video is costly, and semi-supervised learning (SSL) has been studie…
Action RecognitionTemporal Action LocalizationDEARLi: Decoupled Enhancement of Recognition and Localization for Semi-supervised Panoptic Segmentation
Pixel-level annotation is expensive and time-consuming. Semi-supervised segmentation methods address this challenge by learning models on few labeled images alongside a large corpus of unlabeled images. Although foundati…
DecoderGPUPanoptic SegmentationSemantic Segmentation+3