paper-with-me

홈 › Papers

The language of sounds unheard: Exploring musical timbre semantics of large language models

2023-04-16 · Kai Siedenburg, Charalampos Saitis

Semantic dimensions of sound have been playing a central role in understanding the nature of auditory sensory experience as well as the broader relation between perception, language, and meaning. Accordingly, and given the recent proliferation of large language models (LLMs), here we asked whether such models exhibit an organisation of perceptual semantics similar to those observed in humans. Specifically, we prompted ChatGPT, a chatbot based on a state-of-the-art LLM, to rate musical instrument sounds on a set of 20 semantic scales. We elicited multiple responses in separate chats, analogous to having multiple human raters. ChatGPT generated semantic profiles that only partially correlated with human ratings, yet showed robust agreement along well-known psychophysical dimensions of musical sounds such as brightness (bright-dark) and pitch height (deep-high). Exploratory factor analysis suggested the same dimensionality but different spatial configuration of a latent factor space between the chatbot and human ratings. Unexpectedly, the chatbot showed degrees of internal variability that were comparable in magnitude to that of human ratings. Our work highlights the potential of LLMs to capture salient dimensions of human sensory experience.

📄 PDF Abstract BibTeX arXiv:2304.07830

Code (0)

등록된 구현이 없습니다.

Tasks

Chatbot

Similar Papers 제목 키워드 기반

Learning Disentangled Representations of Timbre and Pitch for Musical Instrument Sounds Using Gaussian Mixture Variational Autoencoders

2019-06-19 · Yin-Jyun Luo, Kat Agres, Dorien Herremans

In this paper, we learn disentangled representations of timbre and pitch for musical instrument sounds. We adapt a framework based on variational autoencoders with Gaussian mixture latent distributions. Specifically, we …

Decoder

CAESynth: Real-Time Timbre Interpolation and Pitch Control with Conditional Autoencoders

2021-11-09 · IEEE MLSP 2021 9 · Aaron Valero Puche, Sukhan Lee

In this paper, we present a novel audio synthesizer, CAESynth, based on a conditional autoencoder. CAESynth synthesizes timbre in real-time by interpolating the reference sounds in their shared latent feature space, whil…

Audio SynthesisMixed RealityPitch controlTimbre Interpolation

Contrastive timbre representations for musical instrument and synthesizer retrieval

2025-09-16 · Gwendal Le Vaillant, Yannick Molle arxiv

Efficiently retrieving specific instrument timbres from audio mixtures remains a challenge in digital music production. This paper introduces a contrastive learning framework for musical instrument retrieval, enabling di…

Contrastive LearningData Augmentation

Neural coincidence detection strategies during perception of multi-pitch musical tones

2020-01-17

Multi-pitch perception is investigated in a listening test using 30 recordings of musical sounds with two tones played simultaneously, except for two gong sounds with inharmonic overtone spectrum, judging roughness and s…

DrumGAN: Synthesis of Drum Sounds With Timbral Feature Conditioning Using Generative Adversarial Networks

2020-08-27 · J. Nistal, S. Lattner, G. Richard

Synthetic creation of drum sounds (e.g., in drum machines) is commonly performed using analog or digital synthesis, allowing a musician to sculpt the desired timbre modifying various parameters. Typically, such parameter…

Audio SynthesisGenerative Adversarial Network