paper-with-me

Papers

Resolving Inflectional Ambiguity of Macedonian Adjectives

2022-06-01 · gwll (LREC) 2022 6 · Katerina Zdravkova

Macedonian adjectives are inflected for gender, number, definiteness and degree, with in average 47.98 inflections per headword. The inflection paradigm of qualificative adjectives is even richer, embracing 56.27 morphophonemic alterations. Depending on the word they were derived from, more than 600 Macedonian adjectives have an identical headword and two different word forms for each grammatical category. While non-verbal adjectives alter the root before adding the inflectional suffixes, suffixes of verbal adjectives are added directly to the root. In parallel with the morphological differences, both types of adjectives have a different translation, depending on the category of the words they have been derived from. Nouns that collocate with these adjectives are mutually disjunctive, enabling the resolution of inflectional ambiguity. They are organised as a lexical taxonomy, created using hierarchical divisive clustering. If embedded in the future spell-checking applications, this taxonomy will significantly reduce the risk of forming incorrect inflections, which frequently occur in the daily news and more often in the advertisements and social media.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Extracting inflectional class assignment in Pite Saami: Nouns, verbs and those pesky adjectives

2018-01-01 · WS 2018 1 · Joshua Wilbur

Automatic acquisition of Urdu nouns (along with gender and irregular plurals)

2014-05-01 · LREC 2014 5 · Tafseer Ahmed Khan

The paper describes a set of methods to automatically acquire the Urdu nouns (and its gender) on the basis of inflectional and contextual clues. The algorithms used are a blend of computer{'}s brute force on the corpus a…

Transliteration

TLT-CRF: A Lexicon-supported Morphological Tagger for Latin Based on Conditional Random Fields

2016-05-01 · LREC 2016 5 · Tim vor der Br{\"u}ck, Alex Mehler, er

We present a morphological tagger for Latin, called TTLab Latin Tagger based on Conditional Random Fields (TLT-CRF) which uses a large Latin lexicon. Beyond Part of Speech (PoS), TLT-CRF tags eight inflectional categorie…

POS

Building a Macedonian Recipe Dataset: Collection, Parsing, and Comparative Analysis

2025-10-15 · Darko Sasanski, Dimitar Peshevski, Riste Stojanov, Dimitar Trajanov arxiv

Computational gastronomy increasingly relies on diverse, high-quality recipe datasets to capture regional culinary traditions. Although there are large-scale collections for major languages, Macedonian recipes remain und…

Sentiment Analysis in Twitter for Macedonian

2021-09-27 · RANLP 2015 9 · Dame Jovanoski, Veno Pachovski, Preslav Nakov

We present work on sentiment analysis in Twitter for Macedonian. As this is pioneering work for this combination of language and genre, we created suitable resources for training and evaluating a system for sentiment ana…

Sentiment Analysis