paper-with-me

홈 › Papers

Flexible Japanese Sentence Compression by Relaxing Unit Constraints

2012-12-01 · COLING 2012 12 · Jun Harashima, Sadao Kurohashi
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

SentenceSentence Compression

Similar Papers 제목 키워드 기반

Japanese Sentence Compression with a Large Training Dataset

2017-07-01 · ACL 2017 7 · Shun Hasegawa, Yuta Kikuchi, Hiroya Takamura, Manabu Okumura

In English, high-quality sentence compression models by deleting words have been trained on automatically created large training datasets. We work on Japanese sentence compression by a similar approach. To create a large…

SentenceSentence Compression

A Japanese Word Dependency Corpus

2014-05-01 · LREC 2014 5 · Shinsuke Mori, Hideki Ogura, Tetsuro Sasada

In this paper, we present a corpus annotated with dependency relationships in Japanese. It contains about 30 thousand sentences in various domains. Six domains in Balanced Corpus of Contemporary Written Japanese have par…

ArticlesDependency ParsingMachine TranslationSentence+1

JaParaPat: A Large-Scale Japanese-English Parallel Patent Application Corpus

2025-08-22 · Masaaki Nagata, Katsuki Chousa, Norihito Yasuda arxiv

We constructed JaParaPat (Japanese-English Parallel Patent Application Corpus), a bilingual corpus of more than 300 million Japanese-English sentence pairs from patent applications published in Japan and the United State…

Bilingual Subword Segmentation for Neural Machine Translation

2020-12-01 · COLING 2020 8 · Hiroyuki Deguchi, Masao Utiyama, Akihiro Tamura, Takashi Ninomiya 외

This paper proposed a new subword segmentation method for neural machine translation, {``}Bilingual Subword Segmentation,{''} which tokenizes sentences to minimize the difference between the number of subword units in a …

Machine TranslationSegmentationSentenceTranslation

UD-Japanese BCCWJ: Universal Dependencies Annotation for the Balanced Corpus of Contemporary Written Japanese

2018-11-01 · WS 2018 11 · Mai Omura, Masayuki Asahara

In this paper, we describe a corpus UD Japanese-BCCWJ that was created by converting the Balanced Corpus of Contemporary Written Japanese (BCCWJ), a Japanese language corpus, to adhere to the UD annotation schema. The BC…