paper-with-me

홈 › Papers

Japanese Sentence Compression with a Large Training Dataset

2017-07-01 · ACL 2017 7 · Shun Hasegawa, Yuta Kikuchi, Hiroya Takamura, Manabu Okumura

In English, high-quality sentence compression models by deleting words have been trained on automatically created large training datasets. We work on Japanese sentence compression by a similar approach. To create a large Japanese training dataset, a method of creating English training dataset is modified based on the characteristics of the Japanese language. The created dataset is used to train Japanese sentence compression models based on the recurrent neural network.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

SentenceSentence Compression

Similar Papers 제목 키워드 기반

Flexible Japanese Sentence Compression by Relaxing Unit Constraints

2012-12-01 · COLING 2012 12 · Jun Harashima, Sadao Kurohashi
SentenceSentence Compression

Japanese SimCSE Technical Report

2023-10-30 · Hayato Tsukagoshi, Ryohei Sasano, Koichi Takeda

We report the development of Japanese SimCSE, Japanese sentence embedding models fine-tuned with SimCSE. Since there is a lack of sentence embedding models for Japanese that can be used as a baseline in sentence embeddin…

SentenceSentence EmbeddingSentence-EmbeddingSentence Embeddings

JCSE: Contrastive Learning of Japanese Sentence Embeddings and Its Applications

2023-01-19 · Zihao Chen, Hisashi Handa, Kimiaki Shirahama

Contrastive learning is widely used for sentence representation learning. Despite this prevalence, most studies have focused exclusively on English and few concern domain adaptation for domain-specific downstream tasks, …

Contrastive LearningDomain AdaptationInformation RetrievalLanguage Modelling+7

Domain Adaptation for Japanese Sentence Embeddings with Contrastive Learning based on Synthetic Sentence Generation

2025-03-12 · Zihao Chen, Hisashi Handa, Miho Ohsaki, Kimiaki Shirahama

Several backbone models pre-trained on general domain datasets can encode a sentence into a widely useful embedding. Such sentence embeddings can be further enhanced by domain adaptation that adapts a backbone model to a…

Contrastive LearningDomain AdaptationSemantic Textual SimilaritySentence+2

A Corpus for English-Japanese Multimodal Neural Machine Translation with Comparable Sentences

2020-10-17 · Andrew Merritt, Chenhui Chu, Yuki Arase

Multimodal neural machine translation (NMT) has become an increasingly important area of research over the years because additional modalities, such as image data, can provide more context to textual data. Furthermore, t…

Image CaptioningMachine TranslationNMTSentence+1