paper-with-me

Papers

MultiVitaminBooster at PARSEME Shared Task 2020: Combining Window- and Dependency-Based Features with Multilingual Contextualised Word Embeddings for VMWE Detection

2020-12-01 · COLING (MWE) 2020 12 · Sebastian Gombert, Sabine Bartsch

In this paper, we present MultiVitaminBooster, a system implemented for the PARSEME shared task on semi-supervised identification of verbal multiword expressions - edition 1.2. For our approach, we interpret detecting verbal multiword expressions as a token classification task aiming to decide whether a token is part of a verbal multiword expression or not. For this purpose, we train gradient boosting-based models. We encode tokens as feature vectors combining multilingual contextualized word embeddings provided by the XLM-RoBERTa language model with a more traditional linguistic feature set relying on context windows and dependency relations. Our system was ranked 7th in the official open track ranking of the shared task evaluations with an encoding-related bug distorting the results. For this reason we carry out further unofficial evaluations. Unofficial versions of our systems would have achieved higher ranks.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Language ModelingLanguage Modellingtoken-classificationToken ClassificationWord Embeddings

Similar Papers 제목 키워드 기반

Annotating Verbal MWEs in Irish for the PARSEME Shared Task 1.2

2020-12-01 · COLING (MWE) 2020 12 · Abigail Walsh, Teresa Lynn, Jennifer Foster

This paper describes the creation of two Irish corpora (labelled and unlabelled) for verbal MWEs for inclusion in the PARSEME Shared Task 1.2 on automatic identification of verbal MWEs, and the process of developing verb…

Mumpitz at PARSEME Shared Task 2018: A Bidirectional LSTM for the Identification of Verbal Multiword Expressions

2018-08-01 · COLING 2018 8 · Rafael Ehren, Timm Lichte, Younes Samih

In this paper, we describe Mumpitz, the system we submitted to the PARSEME Shared task on automatic identification of verbal multiword expressions (VMWEs). Mumpitz consists of a Bidirectional Recurrent Neural Network (BR…

Machine TranslationSentenceWord Embeddings

VarIDE at PARSEME Shared Task 2018: Are Variants Really as Alike as Two Peas in a Pod?

2018-08-01 · COLING 2018 8 · Caroline Pasquer, Carlos Ramisch, Agata Savary, Jean-Yves Antoine

We describe the VarIDE system (standing for Variant IDEntification) which participated in the edition 1.1 of the PARSEME shared task on automatic identification of verbal multiword expressions (VMWEs). Our system focuses…

Edition 1.1 of the PARSEME Shared Task on Automatic Identification of Verbal Multiword Expressions

2018-08-01 · COLING 2018 8 · Carlos Ramisch, Silvio Ricardo Cordeiro, Agata Savary, Veronika Vincze 외

This paper describes the PARSEME Shared Task 1.1 on automatic identification of verbal multiword expressions. We present the annotation methodology, focusing on changes from last year{'}s shared task. Novel aspects inclu…

GBD-NER at PARSEME Shared Task 2018: Multi-Word Expression Detection Using Bidirectional Long-Short-Term Memory Networks and Graph-Based Decoding

2018-08-01 · COLING 2018 8 · Tiberiu Boros, Rux Burtica, ra

This paper addresses the issue of multi-word expression (MWE) detection by employing a new decoding strategy inspired after graph-based parsing. We show that this architecture achieves state-of-the-art results with minim…

Feature EngineeringNER