paper-with-me

홈 › Papers

Some clues to build a sound analysis relevant to hearing

2024-01-04 · Laurent Millot

Analysis tools used in research laboratories, for sound synthesis, by musicians or sound engineers can be rather different. Discussion of the assumptions and of the limitations of these tools permits to propose a first tool as relevant and versatile as possible for all the sound actors with a major aim: one must be able to listen to each element of the analysis because hearing is the final reference tool. This tool should also be used, in the future, to reinvestigate the definition of sound (or Acoustics) on the basis of some recent works on musical instrument modeling, speech production and loudspeakers design. Audio illustrations will be given.Paper 6041 presented at the 116th Convention of the Audio Engineering Society, Berlin, 2004

📄 PDF Abstract BibTeX arXiv:2401.02463

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Target Sound Extraction with Variable Cross-modality Clues

2023-03-15 · Chenda Li, Yao Qian, Zhuo Chen, Dongmei Wang 외

Automatic target sound extraction (TSE) is a machine learning approach to mimic the human auditory perception capability of attending to a sound source of interest from a mixture of sources. It often uses a model conditi…

AudioCapsTarget Sound Extraction

Multichannel-to-Multichannel Target Sound Extraction Using Direction and Timestamp Clues

2024-09-19 · Dayun Choi, Jung-Woo Choi

We propose a multichannel-to-multichannel target sound extraction (M2M-TSE) framework for separating multichannel target signals from a multichannel mixture of sound sources. Target sound extraction (TSE) isolates a spec…

Inductive BiasTarget Sound Extraction

SoundBeam: Target sound extraction conditioned on sound-class labels and enrollment clues for increased performance and continuous learning

2022-04-08 · Marc Delcroix, Jorge Bennasar Vázquez, Tsubasa Ochiai, Keisuke Kinoshita 외

In many situations, we would like to hear desired sound events (SEs) while being able to ignore interference. Target sound extraction (TSE) tackles this problem by estimating the audio signal of the sounds of target SE c…

Target Sound Extraction

Object-aware Adaptive-Positivity Learning for Audio-Visual Question Answering

2023-12-20 · Zhangbin Li, Dan Guo, Jinxing Zhou, Jing Zhang 외

This paper focuses on the Audio-Visual Question Answering (AVQA) task that aims to answer questions derived from untrimmed audible videos. To generate accurate answers, an AVQA model is expected to find the most informat…

Audio-visual Question AnsweringAudio-Visual Question Answering (AVQA)Model OptimizationObject+2

Clue-Instruct: Text-Based Clue Generation for Educational Crossword Puzzles

2024-04-09 · Andrea Zugarini, Kamyar Zeinalipour, Surya Sai Kadali, Marco Maggini 외

Crossword puzzles are popular linguistic games often used as tools to engage students in learning. Educational crosswords are characterized by less cryptic and more factual clues that distinguish them from traditional cr…