paper-with-me

홈 › Papers

Sentence-Level Multilingual Multi-modal Embedding for Natural Language Processing

2017-09-01 · RANLP 2017 9 · Iacer Calixto, Qun Liu

We propose a novel discriminative ranking model that learns embeddings from multilingual and multi-modal data, meaning that our model can take advantage of images and descriptions in multiple languages to improve embedding quality. To that end, we introduce an objective function that uses pairwise ranking adapted to the case of three or more input sources. We compare our model against different baselines, and evaluate the robustness of our embeddings on image{--}sentence ranking (ISR), semantic textual similarity (STS), and neural machine translation (NMT). We find that the additional multilingual signals lead to improvements on all three tasks, and we highlight that our model can be used to consistently improve the adequacy of translations generated with NMT models when re-ranking n-best lists.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Machine TranslationNMTRe-RankingSemantic Textual SimilaritySentenceSTSTranslation

Similar Papers 제목 키워드 기반

SAMU-XLSR: Semantically-Aligned Multimodal Utterance-level Cross-Lingual Speech Representation

2022-05-17 · Sameer Khurana, Antoine Laurent, James Glass

We propose the SAMU-XLSR: Semantically-Aligned Multimodal Utterance-level Cross-Lingual Speech Representation learning framework. Unlike previous works on speech representation learning, which learns multilingual context…

Representation LearningRetrievalSentenceSentence Embedding+6

SONAR-SLT: Multilingual Sign Language Translation via Language-Agnostic Sentence Embedding Supervision

2025-10-22 · Yasser Hamidullah, Shakib Yazdani, Cennet Oguz, Josef van Genabith 외 arxiv

Sign language translation (SLT) is typically trained with text in a single spoken language, which limits scalability and cross-language generalization. Earlier approaches have replaced gloss supervision with text-based s…

Sign Language Translation

SONAR: Sentence-Level Multimodal and Language-Agnostic Representations

2023-08-22 · Paul-Ambroise Duquenne, Holger Schwenk, Benoît Sagot

We introduce SONAR, a new multilingual and multimodal fixed-size sentence embedding space. Our single text encoder, covering 200 languages, substantially outperforms existing sentence embeddings such as LASER3 and LabSE …

DecoderMachine TranslationSentenceSentence Embedding+5

FLiP: Towards understanding and interpreting multimodal multilingual sentence embeddings

2026-04-20 · Santosh Kesiraju, Bolaji Yusuf, Šimon Sedláček, Oldřich Plchot 외 arxiv

This paper presents factorized linear projection (FLiP) models for understanding pretrained sentence embedding spaces. We train FLiP models to recover the lexical content from multilingual (LaBSE), multimodal (SONAR) and…

Hierarchical Document Encoder for Parallel Corpus Mining

2019-06-20 · WS 2019 8 · Mandy Guo, Yinfei Yang, Keith Stevens, Daniel Cer 외

We explore using multilingual document embeddings for nearest neighbor mining of parallel data. Three document-level representations are investigated: (i) document embeddings generated by simply averaging multilingual se…

Parallel Corpus MiningSentenceSentence EmbeddingSentence-Embedding+1