paper-with-me

Papers

IndoCollex: A Testbed for Morphological Transformation of Indonesian Word Colloquialism

2021-08-01 · Findings (ACL) 2021 8 · Haryo Akbarianto Wibowo, Made Nindyatama Nityasya, Afra Feyza Akyürek, Suci Fitriany, Alham Fikri Aji, Radityo Eko Prasojo, Derry Tanti Wijaya
📄 PDF Abstract BibTeX

Code (1)

haryoa/indo-collex 공식 구현

Similar Papers 제목 키워드 기반

IDENTIC Corpus: Morphologically Enriched Indonesian-English Parallel Corpus

2012-05-01 · LREC 2012 5 · Septina Dian Larasati

This paper describes the creation process of an Indonesian-English parallel corpus (IDENTIC). The corpus contains 45,000 sentences collected from different sources in different genres. Several manual text preprocessing t…

Spelling Correction

Fine-tuning Pretrained Multilingual BERT Model for Indonesian Aspect-based Sentiment Analysis

2021-03-05 · Annisa Nurul Azhar, Masayu Leylia Khodra

Although previous research on Aspect-based Sentiment Analysis (ABSA) for Indonesian reviews in hotel domain has been conducted using CNN and XGBoost, its model did not generalize well in test data and high number of OOV …

Aspect-Based Sentiment AnalysisAspect-Based Sentiment Analysis (ABSA)Sentiment Analysis

A Neural Network Approach to Create Minangkabau-Indonesia Bilingual Dictionary

2022-06-01 · SIGUL (LREC) 2022 6 · Kartika Resiandi, Yohei Murakami, Arbi Haza Nasution

Indonesia has many varieties of ethnic languages, and most come from the same language family, namely Austronesian languages. Coming from that same language family, the words in Indonesian ethnic languages are very simil…

Decoder

KaWAT: A Word Analogy Task Dataset for Indonesian

2019-06-17 · Kemal Kurniawan

We introduced KaWAT (Kata Word Analogy Task), a new word analogy task dataset for Indonesian. We evaluated on it several existing pretrained Indonesian word embeddings and embeddings trained on Indonesian online news cor…

Word Embeddings

Exploiting Morphological Regularities in Distributional Word Representations

2017-09-01 · EMNLP 2017 9 · Arihant Gupta, Syed Sarfaraz Akhtar, Avijit Vajpayee, Arjit Srivastava 외

We present an unsupervised, language agnostic approach for exploiting morphological regularities present in high dimensional vector spaces. We propose a novel method for generating embeddings of words from their morpholo…

ChunkingDocument ClassificationQuestion AnsweringWord Embeddings