paper-with-me

Papers

LightSAFT: Lightweight Latent Source Aware Frequency Transform for Source Separation

2021-11-24 · Yeong-Seok Jeong, Jinsung Kim, Woosung Choi, Jaehwa Chung, Soonyoung Jung

Conditioned source separations have attracted significant attention because of their flexibility, applicability and extensionality. Their performance was usually inferior to the existing approaches, such as the single source separation model. However, a recently proposed method called LaSAFT-Net has shown that conditioned models can show comparable performance against existing single-source separation models. This paper presents LightSAFT-Net, a lightweight version of LaSAFT-Net. As a baseline, it provided a sufficient SDR performance for comparison during the Music Demixing Challenge at ISMIR 2021. This paper also enhances the existing LightSAFT-Net by replacing the LightSAFT blocks in the encoder with TFC-TDF blocks. Our enhanced LightSAFT-Net outperforms the previous one with fewer parameters.Conditioned source separations have attracted significant attention because of their flexibility, applicability and extensionality. Their performance was usually inferior to the existing approaches, such as the single source separation model. However, a recently proposed method called LaSAFT-Net has shown that conditioned models can show comparable performance against existing single-source separation models. This paper presents LightSAFT-Net, a lightweight version of LaSAFT-Net. As a baseline, it provided a sufficient SDR performance for comparison during the Music Demixing Challenge at ISMIR 2021.

📄 PDF Abstract BibTeX arXiv:2111.12516

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

SpectraFlow: Unifying Structural Pretraining and Frequency Adaptation for Medical Image Segmentation

2026-05-14 · Zhiquan Chen, Haitao Wang, Guowei Zou, Hejun Wu arxiv

Medical image segmentation remains challenging in low-data regimes, where scarce annotations often yield poor generalization and ambiguous boundaries with missing fine structures. Recent self-supervised pretraining has i…

Medical Image Segmentation

FreqMixFormerV2: Lightweight Frequency-aware Mixed Transformer for Human Skeleton Action Recognition

2024-12-29 · Wenhan Wu, Pengfei Wang, Chen Chen, Aidong Lu

Transformer-based human skeleton action recognition has been developed for years. However, the complexity and high parameter count demands of these models hinder their practical applications, especially in resource-const…

Action Recognition

Learning Latent Representations for Image Translation using Frequency Distributed CycleGAN

2025-08-05 · Shivangi Nigam, Adarsh Prasad Behera, Shekhar Verma, P. Nagabhushan arxiv

This paper presents Fd-CycleGAN, an image-to-image (I2I) translation framework that enhances latent representation learning to approximate real data distributions. Building upon the foundation of CycleGAN, our approach i…

Representation LearningStyle Transfer

Towards Efficient Low-rate Image Compression with Frequency-aware Diffusion Prior Refinement

2026-01-15 · Yichong Xia, Yimin Zhou, Jinpeng Wang, Bin Chen arxiv

Recent advancements in diffusion-based generative priors have enabled visually plausible image compression at extremely low bit rates. However, existing approaches suffer from slow sampling processes and suboptimal bit a…

Image ReconstructionImage Compression

DeCo: Frequency-Decoupled Pixel Diffusion for End-to-End Image Generation

2025-11-24 · Zehong Ma, Longhui Wei, Shuai Wang, Shiliang Zhang 외 arxiv

Pixel diffusion aims to generate images directly in pixel space in an end-to-end fashion. This approach avoids the limitations of VAE in the two-stage latent diffusion, offering higher model capacity. Existing pixel diff…

Image Generation