paper-with-me

Papers

Manual Clustering and Spatial Arrangement of Verbs for Multilingual Evaluation and Typology Analysis

2020-12-01 · COLING 2020 8 · Olga Majewska, Ivan Vuli{\'c}, Diana McCarthy, Anna Korhonen

We present the first evaluation of the applicability of a spatial arrangement method (SpAM) to a typologically diverse language sample, and its potential to produce semantic evaluation resources to support multilingual NLP, with a focus on verb semantics. We demonstrate SpAM{'}s utility in allowing for quick bottom-up creation of large-scale evaluation datasets that balance cross-lingual alignment with language specificity. Starting from a shared sample of 825 English verbs, translated into Chinese, Japanese, Finnish, Polish, and Italian, we apply a two-phase annotation process which produces (i) semantic verb classes and (ii) fine-grained similarity scores for nearly 130 thousand verb pairs. We use the two types of verb data to (a) examine cross-lingual similarities and variation, and (b) evaluate the capacity of static and contextualised representation models to accurately reflect verb semantics, contrasting the performance of large language specific pretraining models with their multilingual equivalent on semantic clustering and lexical similarity, across different domains of verb meaning. We release the data from both phases as a large-scale multilingual resource, comprising 85 verb classes and nearly 130k pairwise similarity scores, offering a wealth of possibilities for further evaluation and research on multilingual verb semantics.

📄 PDF Abstract BibTeX

Code (1)

om304/multi-spa-verb 공식 구현

Tasks

ClusteringMultilingual NLPSpecificity

Similar Papers 제목 키워드 기반

Spatial Multi-Arrangement for Clustering and Multi-way Similarity Dataset Construction

2020-05-01 · LREC 2020 5 · Olga Majewska, Diana McCarthy, Jasper van den Bosch, Nikolaus Kriegeskorte 외

We present a novel methodology for fast bottom-up creation of large-scale semantic similarity resources to support development and evaluation of NLP systems. Our work targets verb similarity, but the methodology is equal…

ClusteringSemantic SimilaritySemantic Textual SimilarityWord Similarity

Evaluating Hierarchies of Verb Argument Structure with Hierarchical Clustering

2017-09-01 · EMNLP 2017 9 · Jesse Mu, Joshua K. Hartshorne, Timothy O{'}Donnell

Verbs can only be used with a few specific arrangements of their arguments (syntactic frames). Most theorists note that verbs can be organized into a hierarchy of verb classes based on the frames they admit. Here we show…

ClusteringLanguage AcquisitionNatural Language InferenceSemantic Parsing

Guided Generative Models using Weak Supervision for Detecting Object Spatial Arrangement in Overhead Images

2021-12-10 · Weiwei Duan, Yao-Yi Chiang, Stefan Leyk, Johannes H. Uhl 외

The increasing availability and accessibility of numerous overhead images allows us to estimate and assess the spatial arrangement of groups of geospatial target objects, which can benefit many applications, such as traf…

Decoderobject-detectionObject Detection

Semantic Data Set Construction from Human Clustering and Spatial Arrangement

2021-03-01 · CL (ACL) 2021 3 · Olga Majewska, Diana McCarthy, Jasper J. F. van den Bosch, Nikolaus Kriegeskorte 외

Abstract Research into representation learning models of lexical semantics usually utilizes some form of intrinsic evaluation to ensure that the learned representations reflect human semantic judgments. Lexical semantic …

ClusteringRepresentation LearningSemantic SimilaritySemantic Textual Similarity+1

Detecting Optional Arguments of Verbs

2016-05-01 · LREC 2016 5 · Andr{\'a}s Kornai, D{\'a}vid M{\'a}rk Nemeskey, G{\'a}bor Recski

We propose a novel method for detecting optional arguments of Hungarian verbs using only positive data. We introduce a custom variant of collexeme analysis that explicitly models the noise in verb frames. Our method is, …

Clustering