paper-with-me

Papers

Is Japanese gendered language used on Twitter ? A large scale study

2020-06-29 · Tiziana Carpi, Stefano Maria Iacus

This study analyzes the usage of Japanese gendered language on Twitter. Starting from a collection of 408 million Japanese tweets from 2015 till 2019 and an additional sample of 2355 manually classified Twitter accounts timelines into gender and categories (politicians, musicians, etc). A large scale textual analysis is performed on this corpus to identify and examine sentence-final particles (SFPs) and first-person pronouns appearing in the texts. It turns out that gendered language is in fact used also on Twitter, in about 6% of the tweets, and that the prescriptive classification into "male" and "female" language does not always meet the expectations, with remarkable exceptions. Further, SFPs and pronouns show increasing or decreasing trends, indicating an evolution of the language used on Twitter.

📄 PDF Abstract BibTeX arXiv:2006.15935

Code (0)

등록된 구현이 없습니다.

Tasks

Sentence

Similar Papers 제목 키워드 기반

Morality-based Assertion and Homophily on Social Media: A Cultural Comparison between English and Japanese Languages

2021-08-24 · Maneet Singh, Rishemjit Kaur, Akiko Matsuo, S. R. S. Iyengar 외

Moral psychology is a domain that deals with moral identity, appraisals and emotions. Previous work has primarily focused on moral development and the associated role of culture. Knowing that language is an inherent elem…

Cultural Vocal Bursts Intensity PredictionFairness

Overview of the 2023 ICON Shared Task on Gendered Abuse Detection in Indic Languages

2024-01-08 · Aatman Vaidya, Arnav Arora, Aditya Joshi, Tarunima Prabhakar

This paper reports the findings of the ICON 2023 on Gendered Abuse Detection in Indic Languages. The shared task deals with the detection of gendered abuse in online text. The shared task was conducted as a part of ICON …

Abuse Detection

The Company They Keep: Extracting Japanese Neologisms Using Language Patterns

2018-01-01 · GWC 2018 1 · James Breen, Timothy Baldwin, Francis Bond

We describe an investigation into the identification and extraction of unrecorded potential lexical items in Japanese text by detecting text passages containing selected language patterns typically associated with such i…

Automatically Extracting Variant-Normalization Pairs for Japanese Text Normalization

2017-11-01 · IJCNLP 2017 11 · Itsumi Saito, Kyosuke Nishida, Kugatsu Sadamitsu, Kuniko Saito 외

Social media texts, such as tweets from Twitter, contain many types of non-standard tokens, and the number of normalization approaches for handling such noisy text has been increasing. We present a method for automatical…

Machine TranslationMorphological AnalysisText Normalization

How much is said in a microblog? A multilingual inquiry based on Weibo and Twitter

2015-06-01 · Han-Teng Liao, King-wa Fu, Scott A. Hale

This paper presents a multilingual study on, per single post of microblog text, (a) how much can be said, (b) how much is written in terms of characters and bytes, and (c) how much is said in terms of information content…