paper-with-me

홈 › Papers

LiveBand: Live Accompaniment Generation in the Audio Domain

2026-06-02 · Marco Pasini, Javier Nistal, Ben Hayes, Mathias Rose Bjare, Stefan Lattner, George Fazekas arxiv

We present LiveBand, a real-time system that generates high-fidelity music accompaniments to live audio input, respecting strict causal constraints. Our method trains a causal transformer generator in the continuous latent space of a pre-trained causal audio autoencoder, using adversarial sequence-level supervision from a discriminator. At each timestep, the generator receives only the causally available mix context and Gaussian noise, and predicts accompaniment latents without access to future mix frames or ground-truth target latents. Training is performed in a single parallel forward pass under causal masking, while streaming inference proceeds autoregressively with a rolling attention state. The model's training and inference computations are matched by design, eliminating teacher forcing and the associated exposure bias. On a multi-instrument music accompaniment benchmark, LiveBand improves over prior work on objective measures of audio quality, beat alignment, and mix adherence, while enabling real-time streaming generation without lookahead into the future on consumer hardware.

📄 PDF Abstract BibTeX arXiv:2606.03803

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Towards Real-Time Human-AI Musical Co-Performance: Accompaniment Generation with Latent Diffusion Models and MAX/MSP

2026-04-08 · Tornike Karchkhadze, Shlomo Dubnov arxiv

We present a framework for real-time human-AI musical co-performance, in which a latent diffusion model generates instrumental accompaniment in response to a live stream of context audio. The system combines a MAX/MSP fr…

JukeDrummer: Conditional Beat-aware Audio-domain Drum Accompaniment Generation via Transformer VQ-VAE

2022-10-12 · Yueh-Kao Wu, Ching-Yu Chiu, Yi-Hsuan Yang

This paper proposes a model that generates a drum track in the audio domain to play along to a user-provided drum-free recording. Specifically, using paired data of drumless tracks and the corresponding human-made drum t…

Beat TrackingDecoder

COCOLA: Coherence-Oriented Contrastive Learning of Musical Audio Representations

2024-04-25 · Ruben Ciranni, Giorgio Mariani, Michele Mancusi, Emilian Postolache 외

We present COCOLA (Coherence-Oriented Contrastive Learning for Audio), a contrastive learning method for musical audio representations that captures the harmonic and rhythmic coherence between samples. Our method operate…

Contrastive LearningMusic Generation

Text-to-Song: Towards Controllable Music Generation Incorporating Vocals and Accompaniment

2024-04-14 · Zhiqing Hong, Rongjie Huang, Xize Cheng, Yongqi Wang 외

A song is a combination of singing voice and accompaniment. However, existing works focus on singing voice synthesis and music generation independently. Little attention was paid to explore song synthesis. In this work, …

Music GenerationSinging Voice Synthesis

FastSAG: Towards Fast Non-Autoregressive Singing Accompaniment Generation

2024-05-13 · Jianyi Chen, Wei Xue, Xu Tan, Zhen Ye 외

Singing Accompaniment Generation (SAG), which generates instrumental music to accompany input vocals, is crucial to developing human-AI symbiotic art creation systems. The state-of-the-art method, SingSong, utilizes a mu…

Rhythm