paper-with-me

홈 › Papers

Arrange, Inpaint, and Refine: Steerable Long-term Music Audio Generation and Editing via Content-based Controls

2024-02-14 · Liwei Lin, Gus Xia, Yixiao Zhang, Junyan Jiang

Controllable music generation plays a vital role in human-AI music co-creation. While Large Language Models (LLMs) have shown promise in generating high-quality music, their focus on autoregressive generation limits their utility in music editing tasks. To address this gap, we propose a novel approach leveraging a parameter-efficient heterogeneous adapter combined with a masking training scheme. This approach enables autoregressive language models to seamlessly address music inpainting tasks. Additionally, our method integrates frame-level content-based controls, facilitating track-conditioned music refinement and score-conditioned music arrangement. We apply this method to fine-tune MusicGen, a leading autoregressive music generation model. Our experiments demonstrate promising results across multiple music editing tasks, offering more flexible controls for future AI-driven music editing tools. The source codes and a demo page showcasing our work are available at https://kikyo-16.github.io/AIR.

📄 PDF Abstract BibTeX arXiv:2402.09508

Code (1)

kikyo-16/airgen 공식 구현 pytorch

Tasks

Audio GenerationMusic Generation

Methods 이 논문이 사용한 방법론

Adapter 설명 없음
Focus 설명 없음
Inpainting Train a convolutional neural network to generate the contents of an arbitrary image region conditioned on its surroundings.

Similar Papers 제목 키워드 기반

Short-Term and Long-Term Context Aggregation Network for Video Inpainting

2020-09-12 · ECCV 2020 8 · Ang Li, Shanshan Zhao, Xingjun Ma, Mingming Gong 외

Video inpainting aims to restore missing regions of a video and has many applications such as video editing and object removal. However, existing methods either suffer from inaccurate short-term context aggregation or ra…

Video EditingVideo Inpainting

Fluoroscopic Shape and Pose Tracking of Catheters with Custom Radiopaque Markers

2025-06-11 · Jared Lawson, Rohan Chitale, Nabil Simaan

Safe navigation of steerable and robotic catheters in the cerebral vasculature requires awareness of the catheters shape and pose. Currently, a significant perception burden is placed on interventionalists to mentally re…

NavigatePose EstimationPose Tracking

Urban Architect: Steerable 3D Urban Scene Generation with Layout Prior

2024-04-10 · Fan Lu, Kwan-Yee Lin, Yan Xu, Hongsheng Li 외

Text-to-3D generation has achieved remarkable success via large-scale text-to-image diffusion models. Nevertheless, there is no paradigm for scaling up the methodology to urban scale. Urban scenes, characterized by numer…

3D GenerationModel OptimizationScene GenerationText to 3D

Feature Refinement to Improve High Resolution Image Inpainting

2022-06-27 · Prakhar Kulshreshtha, Brian Pugh, Salma Jiddi

In this paper, we address the problem of degradation in inpainting quality of neural networks operating at high resolutions. Inpainting networks are often unable to generate globally coherent structures at resolutions hi…

Image InpaintingVocal Bursts Intensity Prediction

Towards Interactive Image Inpainting via Sketch Refinement

2023-06-01 · Chang Liu, Shunxin Xu, Jialun Peng, Kaidong Zhang 외

One tough problem of image inpainting is to restore complex structures in the corrupted regions. It motivates interactive image inpainting which leverages additional hints, e.g., sketches, to assist the inpainting proces…

Image Inpainting