A Preliminary Study of Croatian Lexical Substitution
Lexical substitution is a task of determining a meaning-preserving replacement for a word in context. We report on a preliminary study of this task for the Croatian language on a small-scale lexical sample dataset, manually annotated using three different annotation schemes. We compare the annotations, analyze the inter-annotator agreement, and observe a number of interesting language specific details in the obtained lexical substitutes. Furthermore, we apply a recently-proposed, dependency-based lexical substitution model to our dataset. The model achieves a P@3 score of 0.35, which indicates the difficulty of the task.
Code (0)
등록된 구현이 없습니다.
Tasks
Information RetrievalMachine TranslationWord Sense DisambiguationSimilar Papers 제목 키워드 기반
Does Free Word Order Hurt? Assessing the Practical Lexical Function Model for Croatian
The Practical Lexical Function (PLF) model is a model of computational distributional semantics that attempts to strike a balance between expressivity and learnability in predicting phrase meaning and shows competitive r…
Semantic Textual SimilarityA preliminary study of Croatian Language Syllable Networks
This paper presents preliminary results of Croatian syllable networks analysis. Syllable network is a network in which nodes are syllables and links between them are constructed according to their connections within word…
ClusteringDesigning a Croatian Aspectual Derivatives Dictionary: Preliminary Stages
The paper focusses on derivationally connected verbs in Croatian, i.e. on verbs that share the same lexical morpheme and are derived from other verbs via prefixation, suffixation and/or stem alternations. As in other Sla…
CroDeriV: a new resource for processing Croatian morphology
The paper deals with the processing of Croatian morphology and presents CroDeriV ― a newly developed language resource that contains data about morphological structure and derivational relatedness of verbs in Croatian.…
LemmatizationMorphological AnalysisSense-annotating a Lexical Substitution Data Set with Ubyline
We describe the construction of GLASS, a newly sense-annotated version of the German lexical substitution data set used at the GermEval 2015: LexSub shared task. Using the two annotation layers, we conduct the first know…