paper-with-me

Papers

Textual Representations for Crosslingual Information Retrieval

2021-08-01 · ACL (ECNLP) 2021 8 · Hang Zhang, Liling Tan

In this paper, we explored different levels of textual representations for cross-lingual information retrieval. Beyond the traditional token level representation, we adopted the subword and character level representations for information retrieval that had shown to improve neural machine translation by reducing the out-of-vocabulary issues in machine translation. We found that crosslingual information retrieval performance can be improved by combining search results from subwords and token level representation.Additionally, we improved the search performance by combining and re-ranking the result sets from the different text representations for German, French and Japanese.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Cross-Lingual Information RetrievalInformation RetrievalMachine TranslationRe-RankingRetrievalTranslation

Similar Papers 제목 키워드 기반

Low-Resource Parsing with Crosslingual Contextualized Representations

2019-09-19 · CONLL 2019 11 · Phoebe Mulcaire, Jungo Kasai, Noah A. Smith

Despite advances in dependency parsing, languages with small treebanks still present challenges. We assess recent approaches to multilingual contextual word representations (CWRs), and compare them for crosslingual trans…

Dependency Parsing

Crosslingual Transfer Learning for Relation and Event Extraction via Word Category and Class Alignments

2021-11-01 · EMNLP 2021 11 · Minh Van Nguyen, Tuan Ngo Nguyen, Bonan Min, Thien Huu Nguyen

Previous work on crosslingual Relation and Event Extraction (REE) suffers from the monolingual bias issue due to the training of models on only the source language data. An approach to overcome this issue is to use unlab…

Event ExtractionRelationRepresentation LearningTransfer Learning

Fully Unsupervised Crosslingual Semantic Textual Similarity Metric Based on BERT for Identifying Parallel Data

2019-11-01 · CONLL 2019 11 · Chi-kiu Lo, Michel Simard

We present a fully unsupervised crosslingual semantic textual similarity (STS) metric, based on contextual embeddings extracted from BERT {--} Bidirectional Encoder Representations from Transformers (Devlin et al., 2019)…

Machine TranslationNatural Language UnderstandingSemantic Textual SimilaritySTS+1

Distilling Monolingual and Crosslingual Word-in-Context Representations

2024-09-13 · Yuki Arase, Tomoyuki Kajiwara

In this study, we propose a method that distils representations of word meaning in context from a pre-trained masked language model in both monolingual and crosslingual settings. Word representations are the basis for co…

Language ModelingLanguage ModellingSemantic Textual SimilaritySTS

Crosslingual Document Embedding as Reduced-Rank Ridge Regression

2019-04-08 · Martin Josifoski, Ivan S. Paskov, Hristo S. Paskov, Martin Jaggi 외

There has recently been much interest in extending vector-based word representations to multiple languages, such that words can be compared across languages. In this paper, we shift the focus from words to documents and …

Document EmbeddingregressionRetrievalSentence