paper-with-me

홈 › Papers

YouTube AV 50K: An Annotated Corpus for Comments in Autonomous Vehicles

2018-07-30 · Tao Li, Lei Lin, Minsoo Choi, Kaiming Fu, Siyuan Gong, Jian Wang

With one billion monthly viewers, and millions of users discussing and sharing opinions, comments below YouTube videos are rich sources of data for opinion mining and sentiment analysis. We introduce the YouTube AV 50K dataset, a freely-available collections of more than 50,000 YouTube comments and metadata below autonomous vehicle (AV)-related videos. We describe its creation process, its content and data format, and discuss its possible usages. Especially, we do a case study of the first self-driving car fatality to evaluate the dataset, and show how we can use this dataset to better understand public attitudes toward self-driving cars and public reactions to the accident. Future developments of the dataset are also discussed.

📄 PDF Abstract BibTeX arXiv:1807.11227

Code (1)

Eroica-cpp/YouTube-Statistics 공식 구현

Tasks

Autonomous VehiclesOpinion MiningSelf-Driving CarsSentiment Analysis

Similar Papers 제목 키워드 기반

Corpus Creation for Sentiment Analysis in Code-Mixed Tulu Text

2022-06-01 · SIGUL (LREC) 2022 6 · Asha Hegde, Mudoor Devadas Anusha, Sharal Coelho, Hosahalli Lakshmaiah Shashirekha 외

Sentiment Analysis (SA) employing code-mixed data from social media helps in getting insights to the data and decision making for various applications. One such application is to analyze users’ emotions from comments of …

Decision MakingSentiment Analysis

YouDACC: the Youtube Dialectal Arabic Comment Corpus

2014-05-01 · LREC 2014 5 · Ahmed Salama, Houda Bouamor, Behrang Mohit, Kemal Oflazer

This paper presents YOUDACC, an automatically annotated large-scale multi-dialectal Arabic corpus collected from user comments on Youtube videos. Our corpus covers different groups of dialects: Egyptian (EG), Gulf (GU), …

SenTube: A Corpus for Sentiment Analysis on YouTube Social Media

2014-05-01 · LREC 2014 5 · Olga Uryupina, Barbara Plank, Aliaksei Severyn, Agata Rotondi 외

In this paper we present SenTube -- a dataset of user-generated comments on YouTube videos annotated for information content and sentiment polarity. It contains annotations that allow to develop classifiers for several i…

Document ClassificationInformativenessSentiment AnalysisSpam detection+1

Developing a Multilingual Annotated Corpus of Misogyny and Aggression

2020-03-16 · LREC 2020 5 · Shiladitya Bhattacharya, Siddharth Singh, Ritesh Kumar, Akanksha Bansal 외

In this paper, we discuss the development of a multilingual annotated corpus of misogyny and aggression in Indian English, Hindi, and Indian Bangla as part of a project on studying and automatically identifying misogyny …

Matching Theory and Data with Personal-ITY: What a Corpus of Italian YouTube Comments Reveals About Personality

2020-11-11 · COLING (PEOPLES) 2020 12 · Elisa Bassignana, Malvina Nissim, Viviana Patti

As a contribution to personality detection in languages other than English, we rely on distant supervision to create Personal-ITY, a novel corpus of YouTube comments in Italian, where authors are labelled with personalit…