paper-with-me

Papers

Using Social Networks to Improve Language Variety Identification with Neural Networks

2017-11-01 · IJCNLP 2017 11 · Yasuhide Miura, Tomoki Taniguchi, Motoki Taniguchi, Shotaro Misawa, Tomoko Ohkuma

We propose a hierarchical neural network model for language variety identification that integrates information from a social network. Recently, language variety identification has enjoyed heightened popularity as an advanced task of language identification. The proposed model uses additional texts from a social network to improve language variety identification from two perspectives. First, they are used to introduce the effects of homophily. Secondly, they are used as expanded training data for shared layers of the proposed model. By introducing information from social networks, the model improved its accuracy by 1.67-5.56. Compared to state-of-the-art baselines, these improved performances are better in English and comparable in Spanish. Furthermore, we analyzed the cases of Portuguese and Arabic when the model showed weak performances, and found that the effect of homophily is likely to be weak due to sparsity and noises compared to languages with the strong performances.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Language Identification

Similar Papers 제목 키워드 기반

Experiments in Language Variety Geolocation and Dialect Identification

2020-12-01 · VarDial (COLING) 2020 12 · Tommi Jauhiainen, Heidi Jauhiainen, Krister Lindén

In this paper we describe the systems we used when participating in the VarDial Evaluation Campaign organized as part of the 7th workshop on NLP for similar languages, varieties and dialects. The shared tasks we particip…

Dialect Identification

Findings of the VarDial Evaluation Campaign 2021

2021-04-01 · EACL (VarDial) 2021 4 · Bharathi Raja Chakravarthi, Gaman Mihaela, Radu Tudor Ionescu, Heidi Jauhiainen 외

This paper describes the results of the shared tasks organized as part of the VarDial Evaluation Campaign 2021. The campaign was part of the eighth workshop on Natural Language Processing (NLP) for Similar Languages, Var…

Dialect IdentificationLanguage Identification

The performance of multiple language models in identifying offensive language on social media

2023-12-10 · Hao Li, Brandon Bennett

Text classification is an important topic in the field of natural language processing. It has been preliminarily applied in information retrieval, digital library, automatic abstracting, text filtering, word semantic dis…

Information RetrievalRetrievaltext-classificationText Classification

TweetNLP: Cutting-Edge Natural Language Processing for Social Media

2022-06-29 · Jose Camacho-Collados, Kiamehr Rezaee, Talayeh Riahi, Asahi Ushio 외

In this paper we present TweetNLP, an integrated platform for Natural Language Processing (NLP) in social media. TweetNLP supports a diverse set of NLP tasks, including generic focus areas such as sentiment analysis and …

Language IdentificationNamed Entity RecognitionNamed Entity Recognition (NER)Sentiment Analysis

Author Profiling at PAN: from Age and Gender Identification to Language Variety Identification (invited talk)

2017-04-01 · WS 2017 4 · Paolo Rosso

Author profiling is the study of how language is shared by people, a problem of growing importance in applications dealing with security, in order to understand who could be behind an anonymous threat message, and market…

Author ProfilingMarketingSentiment Analysis