paper-with-me

Papers Audio Compression

“Audio Compression” 태그가 달린 논문 42편 · 필터 해제

Language-Codec: Bridging Discrete Codec Representations and Speech Language Models

2024-02-19 · Shengpeng Ji, Minghui Fang, Jialong Zuo, Ziyue Jiang 외

In recent years, large language models have achieved significant success in generative tasks related to speech, audio, music, and other signal domains. A crucial element of these models is the discrete acoustic codecs, w…

Audio CompressionAudio GenerationQuantization

Spiking Music: Audio Compression with Event Based Auto-encoders

2024-02-02 · Martim Lisboa, Guillaume Bellec

Neurons in the brain communicate information via punctual events called spikes. The timing of spikes is thought to carry rich information, but it is not clear how to leverage this in digital systems. We demonstrate that …

Audio CompressionMusic Compression

MAGMA: Music Aligned Generative Motion Autodecoder

2023-09-03 · Sohan Anisetty, Amit Raj, James Hays

Mapping music to dance is a challenging problem that requires spatial and temporal coherence along with a continual synchronization with the music's progression. Taking inspiration from large language models, we introduc…

Audio CompressionDecoderMotion Generation

Edge Storage Management Recipe with Zero-Shot Data Compression for Road Anomaly Detection

2023-07-10 · YeongHyeon Park, UJu Gim, Myung Jin Kim

Recent studies show edge computing-based road anomaly detection systems which may also conduct data collection simultaneously. However, the edge computers will have small data storage but we need to store the collected a…

Anomaly DetectionAudio CompressionAudio Super-ResolutionData Compression+3

Siamese SIREN: Audio Compression with Implicit Neural Representations

2023-06-22 · Luca A. Lanzendörfer, Roger Wattenhofer

Implicit Neural Representations (INRs) have emerged as a promising method for representing diverse data modalities, including 3D shapes, images, and audio. While recent research has demonstrated successful applications o…

Audio Compression

Quantifying Spatial Audio Quality Impairment

2023-06-13 · Karn N. Watcharasupat, Alexander Lerch

Spatial audio quality is a highly multifaceted concept, with many interactions between environmental, geometrical, anatomical, psychological, and contextual considerations. Methods for characterization or evaluation of t…

Audio CompressionMusic Source Separation

High-Fidelity Audio Compression with Improved RVQGAN

2023-06-11 · NeurIPS 2023 11 · Rithesh Kumar, Prem Seetharaman, Alejandro Luebs, Ishaan Kumar 외

Language models have been successfully used to model natural signals, such as images, speech, and music. A key component of these models is a high quality neural compression model that can compress high-dimensional natur…

Audio CompressionAudio GenerationQuantization

DC CoMix TTS: An End-to-End Expressive TTS with Discrete Code Collaborated with Mixer

2023-05-31 · Yerin Choi, Myoung-Wan Koo

Despite the huge successes made in neutral TTS, content-leakage remains a challenge. In this paper, we propose a new input representation and simple architecture to achieve improved prosody modeling. Inspired by the rece…

Audio Compression

Compression with Bayesian Implicit Neural Representations

2023-05-30 · NeurIPS 2023 11 · Zongyu Guo, Gergely Flamich, Jiajun He, Zhibo Chen 외

Many common types of data can be represented as functions that map coordinates to signal values, such as pixel locations to RGB values in the case of an image. Based on this view, data can be compressed by overfitting a …

Audio CompressionQuantization

An investigation of the reconstruction capacity of stacked convolutional autoencoders for log-mel-spectrograms

2023-01-18 · Anastasia Natsiou, Luca Longo, Sean O'Leary

In audio processing applications, the generation of expressive sounds based on high-level representations demonstrates a high demand. These representations can be used to manipulate the timbre and influence the synthesis…

Audio Compression

Combining Automatic Speaker Verification and Prosody Analysis for Synthetic Speech Detection

2022-10-31 · Luigi Attorresi, Davide Salvi, Clara Borrelli, Paolo Bestagini 외

The rapid spread of media content synthesis technology and the potentially damaging impact of audio and video deepfakes on people's lives have raised the need to implement systems able to detect these forgeries automatic…

Audio CompressionFace SwappingRhythmSpeaker Verification+4

High Fidelity Neural Audio Compression

2022-10-24 · Alexandre Défossez, Jade Copet, Gabriel Synnaeve, Yossi Adi

We introduce a state-of-the-art real-time, high-fidelity, audio codec leveraging neural networks. It consists in a streaming encoder-decoder architecture with quantized latent space trained in an end-to-end fashion. We s…

Audio CompressionAudio Signal ProcessingDecoderVocal Bursts Intensity Prediction

Scaling and compressing melodies using geometric similarity measures

2022-09-19 · Luis Evaristo Caraballo, José Miguel Díaz-Báñez, Fabio Rodríguez, Vanesa Sánchez-Canales 외

Melodic similarity measurement is of key importance in music information retrieval. In this paper, we use geometric matching techniques to measure the similarity between two melodies. We represent music as sets of points…

Audio CompressionGeometric MatchingInformation RetrievalMusic Information Retrieval+1

On The Effect Of Coding Artifacts On Acoustic Scene Classification

2021-12-09 · Nagashree K. S. Rao, Nils Peters

Previous DCASE challenges contributed to an increase in the performance of acoustic scene classification systems. State-of-the-art classifiers demand significant processing capabilities and memory which is challenging fo…

Acoustic Scene ClassificationAudio CompressionClassificationScene Classification

Audio Spectral Enhancement: Leveraging Autoencoders for Low Latency Reconstruction of Long, Lossy Audio Sequences

2021-08-08 · Darshan Deshpande, Harshavardhan Abichandani

With active research in audio compression techniques yielding substantial breakthroughs, spectral reconstruction of low-quality audio waves remains a less indulged topic. In this paper, we propose a novel approach for re…

Audio CompressionQuantizationSpectral Reconstruction

UR Channel-Robust Synthetic Speech Detection System for ASVspoof 2021

2021-07-26 · Xinhui Chen, You Zhang, Ge Zhu, Zhiyao Duan

In this paper, we present UR-AIR system submission to the logical access (LA) and the speech deepfake (DF) tracks of the ASVspoof 2021 Challenge. The LA and DF tasks focus on synthetic speech detection (SSD), i.e. detect…

Audio CompressionFace SwappingSynthetic Speech Detectiontext-to-speech+2

Deep Neural Networks and End-to-End Learning for Audio Compression

2021-05-25 · Daniela N. Rim, Inseon Jang, Heeyoul Choi

Recent achievements in end-to-end deep learning have encouraged the exploration of tasks dealing with highly structured data with unified deep network models. Having such models for compressing audio signals has been cha…

Audio CompressionDecoderDeep Learning

ClefNet: Recurrent Autoencoders with Dynamic Time Warping for Near-Lossless Music Compression and Minimal-Latency Transmission

2021-03-15 · Vignav Ramesh, Mason Wang

The onset of coronavirus disease 2019 (COVID-19), an infectious disease caused by severe acute respiratory syndrome coronavirus 2 (SARS-CoV-2), has sparked unprecedented change. Due to the public health guidelines impose…

Audio CompressionDynamic Time WarpingMusic Compression

MP3net: coherent, minute-long music generation from raw audio with a simple convolutional GAN

2021-01-12 · Korneel van den Broek

We present a deep convolutional GAN which leverages techniques from MP3/Vorbis audio compression to produce long, high-quality audio samples with long-range coherence. The model uses a Modified Discrete Cosine Transform …

Audio CompressionMusic Generation

Bayesian Reconstruction of Fourier Pairs

2020-11-09 · Felipe Tobar, Lerko Araya-Hernández, Pablo Huijse, Petar M. Djurić

In a number of data-driven applications such as detection of arrhythmia, interferometry or audio compression, observations are acquired indistinctly in the time or frequency domains: temporal observations allow us to stu…

AstronomyAudio Compression
← 이전 21–40 / 42 다음 →