paper-with-me

홈 › Papers

Facilitating the Manual Annotation of Sounds When Using Large Taxonomies

2018-11-21 · Xavier Favory, Eduardo Fonseca, Frederic Font, Xavier Serra

Properly annotated multimedia content is crucial for supporting advances in many Information Retrieval applications. It enables, for instance, the development of automatic tools for the annotation of large and diverse multimedia collections. In the context of everyday sounds and online collections, the content to describe is very diverse and involves many different types of concepts, often organised in large hierarchical structures called taxonomies. This makes the task of manually annotating content arduous. In this paper, we present our user-centered development of two tools for the manual annotation of audio content from a wide range of types. We conducted a preliminary evaluation of functional prototypes involving real users. The goal is to evaluate them in a real context, engage in discussions with users, and inspire new ideas. A qualitative analysis was carried out including usability questionnaires and semi-structured interviews. This revealed interesting aspects to consider when developing tools for the manual annotation of audio content with labels drawn from large hierarchical taxonomies.

📄 PDF Abstract BibTeX arXiv:1811.10988

Code (0)

등록된 구현이 없습니다.

Tasks

Information RetrievalRetrieval

Similar Papers 제목 키워드 기반

Unsupervised learning human's activities by overexpressed recognized non-speech sounds

2013-11-08 · Serge Smidtas, Magalie Peyrot

Human activity and environment produces sounds such as, at home, the noise produced by water, cough, or television. These sounds can be used to determine the activity in the environment. The objective is to monitor a per…

TAG

The VU Sound Corpus: Adding More Fine-grained Annotations to the Freesound Database

2016-05-01 · LREC 2016 5 · Emiel van Miltenburg, Benjamin Timmermans, Lora Aroyo

This paper presents a collection of annotations (tags or keywords) for a set of 2,133 environmental sounds taken from the Freesound database (www.freesound.org). The annotations are acquired through an open-ended crowd-l…

Recognizing bird species in diverse soundscapes under weak supervision

2021-07-16 · Christof Henkel, Pascal Pfeiffer, Philipp Singer

We present a robust classification approach for avian vocalization in complex and diverse soundscapes, achieving second place in the BirdCLEF2021 challenge. We illustrate how to make full use of pre-trained convolutional…

Robust classification

A dataset for Audio-Visual Sound Event Detection in Movies

2023-02-14 · Rajat Hebbar, Digbalay Bose, Krishna Somandepalli, Veena Vijai 외

Audio event detection is a widely studied audio processing task, with applications ranging from self-driving cars to healthcare. In-the-wild datasets such as Audioset have propelled research in this field. However, many …

Event DetectionSelf-Driving CarsSound ClassificationSound Event Detection

Epic-Sounds: A Large-scale Dataset of Actions That Sound

2023-02-01 · Jaesung Huh, Jacob Chalk, Evangelos Kazakos, Dima Damen 외

We introduce Epic-Sounds, a large-scale dataset of audio annotations capturing temporal extents and class labels within the audio stream of the egocentric videos. We propose an annotation pipeline where annotators tempor…

Action RecognitionSound Classification