paper-with-me

Papers

A Framework for Automatic Acquisition of Croatian and Serbian Verb Aspect from Corpora

2016-05-01 · LREC 2016 5 · Tanja Samard{\v{z}}i{\'c}, Maja Mili{\v{c}}evi{\'c}

Verb aspect is a grammatical and lexical category that encodes temporal unfolding and duration of events described by verbs. It is a potentially interesting source of information for various computational tasks, but has so far not been studied in much depth from the perspective of automatic processing. Slavic languages are particularly interesting in this respect, as they encode aspect through complex and not entirely consistent lexical derivations involving prefixation and suffixation. Focusing on Croatian and Serbian, in this paper we propose a novel framework for automatic classification of their verb types into a number of fine-grained aspectual classes based on the observable morphology of verb forms. In addition, we provide a set of around 2000 verbs classified based on our framework. This set can be used for linguistic research as well as for testing automatic classification on a larger scale. With minor adjustments the approach is also applicable to other Slavic languages

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

General Classification

Similar Papers 제목 키워드 기반

Universal Dependencies for Serbian in Comparison with Croatian and Other Slavic Languages

2017-04-01 · WS 2017 4 · Tanja Samard{\v{z}}i{\'c}, Mirjana Starovi{\'c}, {\v{Z}}eljko Agi{\'c}, Nikola Ljube{\v{s}}i{\'c}

The paper documents the procedure of building a new Universal Dependencies (UDv2) treebank for Serbian starting from an existing Croatian UDv1 treebank and taking into account the other Slavic UD annotation guidelines. W…

Parsing Croatian and Serbian by Using Croatian Dependency Treebanks

2013-10-01 · WS 2013 10 · {\v{Z}}eljko Agi{\'c}, Danijela Merkler, Da{\v{s}}a Berovi{\'c}
Dependency Parsing

New Inflectional Lexicons and Training Corpora for Improved Morphosyntactic Annotation of Croatian and Serbian

2016-05-01 · LREC 2016 5 · Nikola Ljube{\v{s}}i{\'c}, Filip Klubi{\v{c}}ka, {\v{Z}}eljko Agi{\'c}, Ivo-Pavao Jazbec

In this paper we present newly developed inflectional lexcions and manually annotated corpora of Croatian and Serbian. We introduce hrLex and srLex - two freely available inflectional lexicons of Croatian and Serbian - a…

LEMMA

BERTić - The Transformer Language Model for Bosnian, Croatian, Montenegrin and Serbian

2021-04-01 · EACL (BSNLP) 2021 4 · Nikola Ljubešić, Davor Lauc

In this paper we describe a transformer model pre-trained on 8 billion tokens of crawled text from the Croatian, Bosnian, Serbian and Montenegrin web domains. We evaluate the transformer model on the tasks of part-of-spe…

Commonsense Causal ReasoningLanguage ModelingLanguage Modellingnamed-entity-recognition+4

BERTić -- The Transformer Language Model for Bosnian, Croatian, Montenegrin and Serbian

2021-04-19 · Nikola Ljubešić, Davor Lauc

In this paper we describe a transformer model pre-trained on 8 billion tokens of crawled text from the Croatian, Bosnian, Serbian and Montenegrin web domains. We evaluate the transformer model on the tasks of part-of-spe…

Commonsense Causal ReasoningLanguage ModelingLanguage Modellingnamed-entity-recognition+4