paper-with-me

홈 › Papers

Sequence Models for Computational Etymology of Borrowings

2021-08-01 · Findings (ACL) 2021 8 · Winston Wu, Kevin Duh, David Yarowsky
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Tracking Semantic Change in Cognate Sets for English and Romance Languages

2021-08-01 · ACL (LChange) 2021 8 · Ana Sabina Uban, Alina Maria Cristea, Anca Dinu, Liviu P. Dinu 외

Semantic divergence in related languages is a key concern of historical linguistics. We cross-linguistically investigate the semantic divergence of cognate pairs in English and Romance languages, by means of word embeddi…

Word Embeddings

Computational Etymology and Word Emergence

2020-05-01 · LREC 2020 5 · Winston Wu, David Yarowsky

We developed an extensible, comprehensive Wiktionary parser that improves over several existing parsers. We predict the etymology of a word across the full range of etymology types and languages in Wiktionary, showing im…

Overview of ADoBo 2021: Automatic Detection of Unassimilated Borrowings in the Spanish Press

2021-10-29 · Elena Álvarez Mellado, Luis Espinosa Anke, Julio Gonzalo Arroyo, Constantine Lignos 외

This paper summarizes the main findings of the ADoBo 2021 shared task, proposed in the context of IberLef 2021. In this task, we invited participants to detect lexical borrowings (coming mostly from English) in Spanish n…

The power of context: Random Forest classification of near synonyms. A case study in Modern Hindi

2026-04-01 · Jacek Bąkowski arxiv

Synonymy is a widespread yet puzzling linguistic phenomenon. Absolute synonyms theoretically should not exist, as they do not expand language's expressive potential. However, it was suggested that even if synonyms denote…

Semantic Similarity

Detecting Unassimilated Borrowings in Spanish: An Annotated Corpus and Approaches to Modeling

2022-03-30 · ACL 2022 5 · Elena Álvarez-Mellado, Constantine Lignos

This work presents a new resource for borrowing identification and analyzes the performance and errors of several models on this task. We introduce a new annotated corpus of Spanish newswire rich in unassimilated lexical…

Word Embeddings