paper-with-me

홈 › Papers

A Comparison of Word Embeddings for English and Cross-Lingual Chinese Word Sense Disambiguation

2016-11-09 · WS 2016 12 · Hong Jin Kang, Tao Chen, Muthu Kumar Chandrasekaran, Min-Yen Kan

Word embeddings are now ubiquitous forms of word representation in natural language processing. There have been applications of word embeddings for monolingual word sense disambiguation (WSD) in English, but few comparisons have been done. This paper attempts to bridge that gap by examining popular embeddings for the task of monolingual English WSD. Our simplified method leads to comparable state-of-the-art performance without expensive retraining. Cross-Lingual WSD - where the word senses of a word in a source language e come from a separate target translation language f - can also assist in language learning; for example, when providing translations of target vocabulary for learners. Thus we have also applied word embeddings to the novel task of cross-lingual WSD for Chinese and provide a public dataset for further benchmarking. We have also experimented with using word embeddings for LSTM networks and found surprisingly that a basic LSTM network does not work well. We discuss the ramifications of this outcome.

📄 PDF Abstract BibTeX arXiv:1611.02956

Code (0)

등록된 구현이 없습니다.

Tasks

BenchmarkingTranslationWord EmbeddingsWord Sense Disambiguation

Methods 이 논문이 사용한 방법론

Sigmoid Activation 설명 없음
Tanh Activation 설명 없음
LSTM An LSTM is a type of recurrent neural network that addresses the vanishing gradient problem in vanilla…

Similar Papers 제목 키워드 기반

English-Malay Cross-Lingual Embedding Alignment using Bilingual Lexicon Augmentation

2022-05-01 · ACL 2022 5 · Ying Hao Lim, Jasy Suet Yan Liew

As high-quality Malay language resources are still a scarcity, cross lingual word embeddings make it possible for richer English resources to be leveraged for downstream Malay text classification tasks. This paper focuse…

Cross-Lingual Word EmbeddingsMachine Translationtext-classificationText Classification+1

Multilingual Training of Crosslingual Word Embeddings

2017-04-01 · EACL 2017 4 · Long Duong, Hiroshi Kanayama, Tengfei Ma, Steven Bird 외

Crosslingual word embeddings represent lexical items from different languages using the same vector space, enabling crosslingual transfer. Most prior work constructs embeddings for a pair of languages, with English on on…

Bilingual Lexicon InductionDependency ParsingDocument ClassificationGeneral Classification+6

Discovering Lexical Gaps Using Embeddings from Multilingual LLMs

2026-05-23 · Yoonwon Jung, Aaron S. Cohen, Benjamin K. Bergen arxiv

Lexical gaps are words that do not exist in certain languages. They pose challenges for building multilingual lexical resources, for machine translation, and for cross-lingual transfer. Existing lexical gap detection rel…

Cross-Lingual TransferSemantic SimilarityMachine Translation

Identifying Cognates in English-Dutch and French-Dutch by means of Orthographic Information and Cross-lingual Word Embeddings

2020-05-01 · LREC 2020 5 · Els Lefever, Sofie Labat, Pranaydeep Singh

This paper investigates the validity of combining more traditional orthographic information with cross-lingual word embeddings to identify cognate pairs in English-Dutch and French-Dutch. In a first step, lists of potent…

Cross-Lingual Word EmbeddingsWord Embeddings

Double Trouble: Bilingual Pretraining Leaves Language-Conditioned Effects in Shared-Language Representations

2026-08-27 · Anjishnu Mukherjee, Ziwei Zhu, Antonios Anastasopoulos arxiv

A concept can carry different associations across languages, while modern language models learn English alongside many other languages during pretraining. Yet comparisons among existing models cannot easily isolate how a…