Crowdsourcing High-Quality Parallel Data Extraction from Twitter
Code (0)
등록된 구현이 없습니다.
Tasks
Machine TranslationVocal Bursts Intensity PredictionWord Sense DisambiguationSimilar Papers 제목 키워드 기반
Milimili. Collecting Parallel Data via Crowdsourcing
We present a methodology for gathering a parallel corpus through crowdsourcing, which is more cost-effective than hiring professional translators, albeit at the expense of quality. Additionally, we have made available ex…
A Japanese-Chinese Parallel Corpus Using Crowdsourcing for Web Mining
Using crowdsourcing, we collected more than 10,000 URL pairs (parallel top page pairs) of bilingual websites that contain parallel documents and created a Japanese-Chinese parallel corpus of 4.6M sentence pairs from thes…
SentenceTranslationWord TranslationA Meta-framework for Spatiotemporal Quantity Extraction from Text
News events are often associated with quantities (e.g., the number of COVID-19 patients or the number of arrests in a protest), and it is often important to extract their type, time, and location from unstructured text i…
A Meta-framework for Spatiotemporal Quantity Extraction from Text
News events are often associated with quantities (e.g., the number of COVID-19 patients or the number of arrests in a protest), and it is often important to extract their type, time, and location from unstructured text i…
Crowdsourcing for Evaluating Machine Translation Quality
The recent popularity of machine translation has increased the demand for the evaluation of translations. However, the traditional evaluation approach, manual checking by a bilingual professional, is too expensive and to…
Machine TranslationSentenceTranslation