paper-with-me

홈 › Papers

The POTUS Corpus, a Database of Weekly Addresses for the Study of Stance in Politics and Virtual Agents

2020-05-01 · LREC 2020 5 · Thomas Janssoone, K{\'e}vin Bailly, Ga{\"e}l Richard, Chlo{\'e} Clavel

One of the main challenges in the field of Embodied Conversational Agent (ECA) is to generate socially believable agents. The common strategy for agent behaviour synthesis is to rely on dedicated corpus analysis. Such a corpus is composed of multimedia files of socio-emotional behaviors which have been annotated by external observers. The underlying idea is to identify interaction information for the agent{'}s socio-emotional behavior by checking whether the intended socio-emotional behavior is actually perceived by humans. Then, the annotations can be used as learning classes for machine learning algorithms applied to the social signals. This paper introduces the POTUS Corpus composed of high-quality audio-video files of political addresses to the American people. Two protagonists are present in this database. First, it includes speeches of former president Barack Obama to the American people. Secondly, it provides videos of these same speeches given by a virtual agent named Rodrigue. The ECA reproduces the original address as closely as possible using social signals automatically extracted from the original one. Both are annotated for social attitudes, providing information about the stance observed in each file. It also provides the social signals automatically extracted from Obama{'}s addresses used to generate Rodrigue{'}s ones.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Methods 이 논문이 사용한 방법론

American 설명 없음

Similar Papers 제목 키워드 기반

LLM-POTUS Score: A Framework of Analyzing Presidential Debates with Large Language Models

2024-09-12 · Zhengliang Liu, Yiwei Li, Oleksandra Zolotarevych, Rongwei Yang 외

Large language models have demonstrated remarkable capabilities in natural language processing, yet their application to political discourse analysis remains underexplored. This paper introduces a novel approach to evalu…

Merkel Podcast Corpus: A Multimodal Dataset Compiled from 16 Years of Angela Merkel's Weekly Video Podcasts

2022-05-24 · Debjoy Saha, Shravan Nayak, Timo Baumann

We introduce the Merkel Podcast Corpus, an audio-visual-text corpus in German collected from 16 years of (almost) weekly Internet podcasts of former German chancellor Angela Merkel. To the best of our knowledge, this is …

Face DetectionFace GenerationSpeaker RecognitionTalking Face Generation

Merkel Podcast Corpus: A Multimodal Dataset Compiled from 16 Years of Angela Merkel’s Weekly Video Podcasts

2022-06-01 · LREC 2022 6 · Debjoy Saha, Shravan Nayak, Timo Baumann

We introduce the Merkel Podcast Corpus, an audio-visual-text corpus in German collected from 16 years of (almost) weekly Internet podcasts of former German chancellor Angela Merkel. To the best of our knowledge, this is …

Face DetectionFace GenerationSpeaker RecognitionTalking Face Generation

Corpus for Children's Writing with Enhanced Output for Specific Spelling Patterns (2nd and 3rd Grade)

2016-05-01 · LREC 2016 5 · Kay Berkling

This paper describes the collection of the H1 Corpus of children{'}s weekly writing over the course of 3 months in 2nd and 3rd grades, aged 7-11. The texts were collected within the normal classroom setting by the teache…

A Grammar-informed Corpus-based Sentence Database for Linguistic and Computational Studies

2012-05-01 · LREC 2012 5 · Hongzhi Xu, Helen Kai-yun Chen, Chu-Ren Huang, Qin Lu 외

We adopt the corpus-informed approach to example sentence selections for the construction of a reference grammar. In the process, a database containing sentences that are carefully selected by linguistic experts includin…

Chinese Word SegmentationPOSPOS TaggingSentence