MIPA: Mutual Information Based Paraphrase Acquisition via Bilingual Pivoting
We present a pointwise mutual information (PMI)-based approach to formalize paraphrasability and propose a variant of PMI, called MIPA, for the paraphrase acquisition. Our paraphrase acquisition method first acquires lexical paraphrase pairs by bilingual pivoting and then reranks them by PMI and distributional similarity. The complementary nature of information from bilingual corpora and from monolingual corpora makes the proposed method robust. Experimental results show that the proposed method substantially outperforms bilingual pivoting and distributional similarity themselves in terms of metrics such as MRR, MAP, coverage, and Spearman{'}s correlation.
Code (1)
Tasks
Learning Word EmbeddingsSemantic Textual SimilarityWord AlignmentWord EmbeddingsSimilar Papers 제목 키워드 기반
A Declarative-Procedural Perspective on Expert Routing in Bilingual Mixture-of-Experts Language Models
We investigate whether Mixture-of-Experts (MoE) language models develop linguistically structured expert routing during bilingual language acquisition. Inspired by the Declarative-Procedural framework, we analyze lexical…
Language AcquisitionParaphrase Acquisition from Image Captions
We propose to use image captions from the Web as a previously underutilized resource for paraphrases (i.e., texts with the same "message") and to create and analyze a corresponding dataset. When an image is reused on the…
ArticlesImage CaptioningParaphrasing Revisited with Neural Machine Translation
Recognizing and generating paraphrases is an important component in many natural language processing applications. A well-established technique for automatically extracting paraphrases leverages bilingual corpora to find…
Machine TranslationQuestion AnsweringSemantic ParsingSemantic Role Labeling+1Optimizing rgb-d semantic segmentation through multi-modal interaction and pooling attention
Semantic segmentation of RGB-D images involves understanding the appearance and spatial relationships of objects within a scene, which requires careful consideration of various factors. However, in indoor environments, t…
DecoderSegmentationSemantic Segmentation