paper-with-me

홈 › Papers

The CIRDO Corpus: Comprehensive Audio/Video Database of Domestic Falls of Elderly People

2016-05-01 · LREC 2016 5 · Michel Vacher, Sa{\"\i}da Bouakaz, Marc-Eric Bobillier Chaumon, Fr{\'e}d{\'e}ric Aman, R. A. Khan, Slima Bekkadja, Fran{\c{c}}ois Portet, Erwan Guillou, Solange Rossato, Benjamin Lecouteux

Ambient Assisted Living aims at enhancing the quality of life of older and disabled people at home thanks to Smart Homes. In particular, regarding elderly living alone at home, the detection of distress situation after a fall is very important to reassure this kind of population. However, many studies do not include tests in real settings, because data collection in this domain is very expensive and challenging and because of the few available data sets. The C IRDO corpus is a dataset recorded in realistic conditions in D OMUS , a fully equipped Smart Home with microphones and home automation sensors, in which participants performed scenarios including real falls on a carpet and calls for help. These scenarios were elaborated thanks to a field study involving elderly persons. Experiments related in a first part to distress detection in real-time using audio and speech analysis and in a second part to fall detection using video analysis are presented. Results show the difficulty of the task. The database can be used as standardized database by researchers to evaluate and compare their systems for elderly person{'}s assistance.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

The AV-LASYN Database : A synchronous corpus of audio and 3D facial marker data for audio-visual laughter synthesis

2014-05-01 · LREC 2014 5 · H{\"u}seyin {\c{C}}akmak, J{\'e}r{\^o}me Urbain, Thierry Dutoit, Jo{\"e}lle Tilmanne

A synchronous database of acoustic and 3D facial marker data was built for audio-visual laughter synthesis. Since the aim is to use this database for HMM-based modeling and synthesis, the amount of collected data from on…

Dimensionality ReductionSpeech Synthesis

SUTAV: A Turkish Audio-Visual Database

2012-05-01 · LREC 2012 5 · Ibrahim Saygin Topkaya, Hakan Erdogan

This paper contains information about the ''''''``Sabanci University Turkish Audio-Visual (SUTAV)'''''''' database. The main aim of collecting SUTAV database was to obtain a large audio-visual collection of spoken words,…

Audio-Visual Speech RecognitionPerson Identificationspeech-recognitionSpeech Recognition+1

How Does Audio Influence Visual Attention in Omnidirectional Videos? Database and Model

2024-08-10 · Yuxin Zhu, Huiyu Duan, Kaiwei Zhang, Yucheng Zhu 외

Understanding and predicting viewer attention in omnidirectional videos (ODVs) is crucial for enhancing user engagement in virtual and augmented reality applications. Although both audio and visual modalities are essenti…

PredictionSaliency Prediction

Subjective and Objective Audio-Visual Quality Assessment for User Generated Content

2023-07-10 · IEEE Transactions on Image Processing 2023 7 · Yuqin Cao, Xiongkuo Min, Wei Sun, Guangtao Zhai

In recent years, User Generated Content (UGC) has grown dramatically in video sharing applications. It is necessary for service-providers to use video quality assessment (VQA) to monitor and control users’ Quality of Exp…

Video Quality AssessmentVisual Question Answering (VQA)

CirdoX: an on/off-line multisource speech and sound analysis software

2016-05-01 · LREC 2016 5 · Fr{\'e}d{\'e}ric Aman, Michel Vacher, Fran{\c{c}}ois Portet, William Duclot 외

Vocal User Interfaces in domestic environments recently gained interest in the speech processing community. This interest is due to the opportunity of using it in the framework of Ambient Assisted Living both for home au…

General Classification