paper-with-me

Papers

HyperSound: Generating Implicit Neural Representations of Audio Signals with Hypernetworks

2022-11-03 · Filip Szatkowski, Karol J. Piczak, Przemysław Spurek, Jacek Tabor, Tomasz Trzciński

Implicit neural representations (INRs) are a rapidly growing research field, which provides alternative ways to represent multimedia signals. Recent applications of INRs include image super-resolution, compression of high-dimensional signals, or 3D rendering. However, these solutions usually focus on visual data, and adapting them to the audio domain is not trivial. Moreover, it requires a separately trained model for every data sample. To address this limitation, we propose HyperSound, a meta-learning method leveraging hypernetworks to produce INRs for audio signals unseen at training time. We show that our approach can reconstruct sound waves with quality comparable to other state-of-the-art models.

📄 PDF Abstract BibTeX arXiv:2211.01839

Code (0)

등록된 구현이 없습니다.

Tasks

Image Super-ResolutionMeta-LearningSuper-Resolution

Similar Papers 제목 키워드 기반

Hypernetworks build Implicit Neural Representations of Sounds

2023-02-09 · Filip Szatkowski, Karol J. Piczak, Przemysław Spurek, Jacek Tabor 외

Implicit Neural Representations (INRs) are nowadays used to represent multimedia signals across various real-life applications, including image super-resolution, image compression, or 3D rendering. Existing methods that …

Image CompressionImage Super-ResolutionMeta-LearningSuper-Resolution

A Hypernetwork-Based Approach to KAN Representation of Audio Signals

2025-03-04 · Patryk Marszałek, Maciej Rut, Piotr Kawa, Przemysław Spurek 외

Implicit neural representations (INR) have gained prominence for efficiently encoding multimedia data, yet their applications in audio signals remain limited. This study introduces the Kolmogorov-Arnold Network (KAN), a …

AD-NeRF: Audio Driven Neural Radiance Fields for Talking Head Synthesis

2021-03-20 · ICCV 2021 10 · Yudong Guo, Keyu Chen, Sen Liang, Yong-Jin Liu 외

Generating high-fidelity talking head video by fitting with the input audio sequence is a challenging problem that receives considerable attentions recently. In this paper, we address this problem with the aid of neural …

NeRFTalking Face Generation

Sounding Video Generator: A Unified Framework for Text-guided Sounding Video Generation

2023-03-29 · Jiawei Liu, Weining Wang, Sihan Chen, Xinxin Zhu 외

As a combination of visual and audio signals, video is inherently multi-modal. However, existing video generation methods are primarily intended for the synthesis of visual frames, whereas audio signals in realistic vide…

Audio GenerationContrastive LearningDecoderVideo Generation

“Style” Transfer for Musical Audio Using Multiple Time-Frequency Representations

2018-01-01 · ICLR 2018 1 · Shaun Barry, Youngmoo Kim

Neural Style Transfer has become a popular technique for generating images of distinct artistic styles using convolutional neural networks. This recent success in image style transfer has raised the question of whether s…

Style TransferTexture Synthesis