paper-with-me

Papers

GANStrument: Adversarial Instrument Sound Synthesis with Pitch-invariant Instance Conditioning

2022-11-10 · Gaku Narita, Junichi Shimizu, Taketo Akama

We propose GANStrument, a generative adversarial model for instrument sound synthesis. Given a one-shot sound as input, it is able to generate pitched instrument sounds that reflect the timbre of the input within an interactive time. By exploiting instance conditioning, GANStrument achieves better fidelity and diversity of synthesized sounds and generalization ability to various inputs. In addition, we introduce an adversarial training scheme for a pitch-invariant feature extractor that significantly improves the pitch accuracy and timbre consistency. Experimental results show that GANStrument outperforms strong baselines that do not use instance conditioning in terms of generation quality and input editability. Qualitative examples are available online.

📄 PDF Abstract BibTeX arXiv:2211.05385

Code (0)

등록된 구현이 없습니다.

Tasks

Diversity

Similar Papers 제목 키워드 기반

HyperGANStrument: Instrument Sound Synthesis and Editing with Pitch-Invariant Hypernetworks

2024-01-09 · Zhe Zhang, Taketo Akama

GANStrument, exploiting GANs with a pitch-invariant feature extractor and instance conditioning technique, has shown remarkable capabilities in synthesizing realistic instrument sounds. To further improve the reconstruct…

Diversity

Signal Representations for Synthesizing Audio Textures with Generative Adversarial Networks

2021-03-12 · Chitralekha Gupta, Purnima Kamath, Lonce Wyse

Generative Adversarial Networks (GANs) currently achieve the state-of-the-art sound synthesis quality for pitched musical instruments using a 2-channel spectrogram representation consisting of log magnitude and instantan…

Audio Synthesis

Learning Disentangled Representations of Timbre and Pitch for Musical Instrument Sounds Using Gaussian Mixture Variational Autoencoders

2019-06-19 · Yin-Jyun Luo, Kat Agres, Dorien Herremans

In this paper, we learn disentangled representations of timbre and pitch for musical instrument sounds. We adapt a framework based on variational autoencoders with Gaussian mixture latent distributions. Specifically, we …

Decoder

A Unified Model for Zero-shot Music Source Separation, Transcription and Synthesis

2021-08-07 · Liwei Lin, Qiuqiang Kong, Junyan Jiang, Gus Xia

We propose a unified model for three inter-related tasks: 1) to \textit{separate} individual sound sources from a mixed music audio, 2) to \textit{transcribe} each sound source to MIDI notes, and 3) to\textit{ synthesize…

DecoderDisentanglementMusic Source SeparationMusic Transcription+1

Pitch-Conditioned Instrument Sound Synthesis From an Interactive Timbre Latent Space

2025-10-05 · Christian Limberg, Fares Schulz, Zhe Zhang, Stefan Weinzierl arxiv

This paper presents a novel approach to neural instrument sound synthesis using a two-stage semi-supervised learning framework capable of generating pitch-accurate, high-quality music samples from an expressive timbre la…

Audio Generation