Automatically Annotating A Five-Billion-Word Corpus of Japanese Blogs for Affect and Sentiment Analysis
Code (0)
등록된 구현이 없습니다.
Tasks
Opinion MiningSentiment AnalysisSimilar Papers 제목 키워드 기반
1.5 billion words Arabic Corpus
This study is an attempt to build a contemporary linguistic corpus for Arabic language. The corpus produced, is a text corpus includes more than five million newspaper articles. It contains over a billion and a half word…
ArticlesAnnotating the MASC Corpus with BabelNet
In this paper we tackle the problem of automatically annotating, with both word senses and named entities, the MASC 3.0 corpus, a large English corpus covering a wide range of genres of written and spoken text. We use Ba…
Entity LinkingReading ComprehensionRelation ExtractionWord Sense DisambiguationMultilingual Named Entity Recognition for Medieval Charters Using Stacked Embeddings and Bert-based Models.
In recent years the availability of medieval charter texts has increased thanks to advances in OCR and HTR techniques. But the lack of models that automatically structure the textual output continues to hinder the extrac…
HTRMultilingual Named Entity Recognitionnamed-entity-recognitionNamed Entity Recognition+2Shamela: A Large-Scale Historical Arabic Corpus
Arabic is a widely-spoken language with a rich and long history spanning more than fourteen centuries. Yet existing Arabic corpora largely focus on the modern period or lack sufficient diachronic information. We develop …
Interannotator Agreement for Lexico-Semantic Annotation of a Corpus
This paper examines the procedure for lexico-semantic annotation of the Basic Corpus of Polish Metaphors that is the first step for annotating metaphoric expressions occurring in it. The procedure involves correcting the…