Identifying collocations using cross-lingual association measures
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
A comparison of statistical association measures for identifying dependency-based collocations in various languages.
This paper presents an exploration of different statistical association measures to automatically identify collocations from corpora in English, Portuguese, and Spanish. To evaluate the impact of the association metrics …
Identifying Phrasemes via Interlingual Association Measures -- A Data-driven Approach on Dependency-parsed and Word-aligned Parallel Corpora
This is a preprint of the article "Identifying Phrasemes via Interlingual Association Measures" that was presented in February 2016 at the LeKo (Lexical combinations and typified speech in a multilingual context) confere…
All That Glitters is Not Gold: A Gold Standard of Adjective-Noun Collocations for German
In this paper we present the GerCo dataset of adjective-noun collocations for German, such as alter Freund {`}old friend{'} and tiefe Liebe {`}deep love{'}. The annotation has been performed by experts based on the annot…
AllWord EmbeddingsCollocation or Free Combination? --- Applying Machine Translation Techniques to identify collocations in Japanese
This work presents an initial investigation on how to distinguish collocations from free combinations. The assumption is that, while free combinations can be literally translated, the overall meaning of collocations is d…
Machine TranslationTranslationUsing bilingual word-embeddings for multilingual collocation extraction
This paper presents a new strategy for multilingual collocation extraction which takes advantage of parallel corpora to learn bilingual word-embeddings. Monolingual collocation candidates are retrieved using Universal De…
Machine TranslationTranslationWord Embeddings