paper-with-me

홈 › Papers

Continuous Autoregressive Models with Noise Augmentation Avoid Error Accumulation

2024-11-27 · Marco Pasini, Javier Nistal, Stefan Lattner, George Fazekas

Autoregressive models are typically applied to sequences of discrete tokens, but recent research indicates that generating sequences of continuous embeddings in an autoregressive manner is also feasible. However, such Continuous Autoregressive Models (CAMs) can suffer from a decline in generation quality over extended sequences due to error accumulation during inference. We introduce a novel method to address this issue by injecting random noise into the input embeddings during training. This procedure makes the model robust against varying error levels at inference. We further reduce error accumulation through an inference procedure that introduces low-level noise. Experiments on musical audio generation show that CAM substantially outperforms existing autoregressive and non-autoregressive approaches while preserving audio quality over extended sequences. This work paves the way for generating continuous embeddings in a purely autoregressive setting, opening new possibilities for real-time and interactive generative applications.

📄 PDF Abstract BibTeX arXiv:2411.18447

Code (0)

등록된 구현이 없습니다.

Tasks

Audio Generation

Methods 이 논문이 사용한 방법론

CAM Class activation maps could be used to interpret the prediction decision made by the convolutional neural network (CNN). Image source: [Learning Deep Features for…

Similar Papers 제목 키워드 기반

Autoregression-Free Neural Operators for Time-Dependent PDEs

2026-05-25 · Jiaquan Zhang, Caiyan Qin, Haoyu Bian, Libin Cai 외 arxiv

Neural operators learn mappings from function-dependent inputs to solutions, providing an effective framework for solving partial differential equations (PDEs). For time-dependent PDEs, existing methods typically perform…

AVATAR: Robust Voice Search Engine Leveraging Autoregressive Document Retrieval and Contrastive Learning

2023-09-04 · Yi-Cheng Wang, Tzu-Ting Yang, Hsin-Wei Wang, Bi-Cheng Yan 외

Voice, as input, has progressively become popular on mobiles and seems to transcend almost entirely text input. Through voice, the voice search (VS) system can provide a more natural way to meet user's information needs.…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)Contrastive LearningOpen-Domain Question Answering+4

Efficient Speech Language Modeling via Energy Distance in Continuous Latent Space

2025-05-19 · Zhengrui Ma, Yang Feng, Chenze Shao, Fandong Meng 외

We introduce SLED, an alternative approach to speech language modeling by encoding speech waveforms into sequences of continuous latent representations and modeling them autoregressively using an energy distance objectiv…

Language ModelingLanguage ModellingQuantizationSpeech Synthesis

Variable Skipping for Autoregressive Range Density Estimation

2020-07-10 · ICML 2020 1 · Eric Liang, Zongheng Yang, Ion Stoica, Pieter Abbeel 외

Deep autoregressive models compute point likelihood estimates of individual data points. However, many applications (i.e., database cardinality estimation) require estimating range densities, a capability that is under-e…

Data AugmentationDensity Estimation

DELTA: Language Diffusion-based EEG-to-Text Architecture

2025-11-22 · Mingyu Jeon, Hyobin Kim arxiv

Electroencephalogram (EEG)-to-text remains challenging due to high-dimensional noise, subject variability, and error accumulation in autoregressive decoding. We introduce DELTA, which pairs a Residual Vector Quantization…

Text Generation