paper-with-me

홈 › Papers

Learning non-concatenative morphology

2013-08-01 · WS 2013 8 · Michelle Fullwood, Tim O{'}Donnell
📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Bayesian Inference

Similar Papers 제목 키워드 기반

Adaptor Grammars for Learning Non-Concatenative Morphology

2013-10-01 · EMNLP 2013 10 · Jan A. Botha, Phil Blunsom
Information RetrievalLanguage AcquisitionMachine Translation

Unsupervised acquisition of concatenative morphology

2012-05-01 · LREC 2012 5 · Lionel Nicolas, Jacques Farr{\'e}, C{\'e}cile Darme

Among the linguistic resources formalizing a language, morphological rules are among those that can be achieved in a reasonable time. Nevertheless, since the construction of such resource can require linguistic expertise…

Morphological Inflection

How Suitable Are Subword Segmentation Strategies for Translating Non-Concatenative Morphology?

2021-09-02 · Findings (EMNLP) 2021 11 · Chantal Amrhein, Rico Sennrich

Data-driven subword segmentation has become the default strategy for open-vocabulary machine translation and other NLP tasks, but may not be sufficiently generic for optimal learning of non-concatenative morphology. We d…

Machine TranslationSegmentationTranslation

Splintering Nonconcatenative Languages for Better Tokenization

2025-03-18 · Bar Gazit, Shaltiel Shmidman, Avi Shmidman, Yuval Pinter

Common subword tokenization algorithms like BPE and UnigramLM assume that text can be split into meaningful units by concatenative measures alone. This is not true for languages such as Hebrew and Arabic, where morpholog…

Constrained Sequence-to-sequence Semitic Root Extraction for Enriching Word Embeddings

2019-08-01 · WS 2019 8 · Ahmed El-Kishky, Xingyu Fu, Aseel Addawood, Nahil Sobh 외

In this paper, we tackle the problem of {``}root extraction{''} from words in the Semitic language family. A challenge in applying natural language processing techniques to these languages is the data sparsity problem th…

Language ModelingLanguage ModellingWord EmbeddingsWord Similarity