Lexical Comparison Between Wikipedia and Twitter Corpora by Using Word Embeddings
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationWord EmbeddingsSimilar Papers 제목 키워드 기반
Wikification for Scriptio Continua
The fact that Japanese employs scriptio continua, or a writing system without spaces, complicates the first step of an NLP pipeline. Word segmentation is widely used in Japanese language processing, and lexical knowledge…
SegmentationVMWE discovery: a comparative analysis between Literature and Twitter Corpora
We evaluate manually five lexical association measurements as regards the discovery of Modern Greek verb multiword expressions with two or more lexicalised components usingmwetoolkit3 (Ramisch et al., 2010). We use Twitt…
Using four different online media sources to forecast the crude oil price
This study looks for signals of economic awareness on online social media and tests their significance in economic predictions. The study analyses, over a period of two years, the relationship between the West Texas Inte…
ArticlesTwitter as a Lifeline: Human-annotated Twitter Corpora for NLP of Crisis-related Messages
Microblogging platforms such as Twitter provide active communication channels during mass convergence and emergency events such as earthquakes, typhoons. During the sudden onset of a crisis situation, affected people pos…
Disaster ResponseHumanitarianWord EmbeddingsCultural Topic Modelling over Novel Wikipedia Corpora for South-Slavic Languages
There is a shortage of high-quality corpora for South-Slavic languages. Such corpora are useful to computer scientists and researchers in social sciences and humanities alike, focusing on numerous linguistic, content ana…
Cultural Vocal Bursts Intensity Prediction