paper-with-me

홈 › Papers

Modality annotation for Portuguese: from manual annotation to automatic labeling

2016-09-01 · LILT 2016 9 · Amália Mendes, Iris Hendrickx, Liciana Ávila, Paulo Quaresma, Teresa Gonҫalves, João Sequeira

We investigate modality in Portuguese and we combine a linguistic perspective with an application-oriented perspective on modality. We design an annotation scheme reflecting theoretical linguistic concepts and apply this schema to a small corpus sample to show how the scheme deals with real world language usage. We present two schemas for Portuguese, one for spoken Brazilian Portuguese and one for written European Portuguese. Furthermore, we use the annotated data not only to study the linguistic phenomena of modality, but also to train a practical text mining tool to detect modality in text automatically. The modality tagger uses a machine learning classifier trained on automatically extracted features from a syntactic parser. As we only have a small annotated sample available, the tagger was evaluated on 11 modal verbs that are frequent in our corpus and that denote more than one modal meaning. Finally, we discuss several valuable insights into the complexity of the semantic concept of modality that derive from the process of manual annotation of the corpus and from the analysis of the results of the automatic labeling: ambiguity and the semantic and syntactic properties typically associated to one modal meaning in context, and also the interaction of modality with negation and focus. The knowledge gained from the manual annotation task leads us to propose a new unified scheme for modality that applies to the two Portuguese varieties and covers both written and spoken data.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Negation

Similar Papers 제목 키워드 기반

Towards a Unified Approach to Modality Annotation in Portuguese

2015-04-01 · WS 2015 4 · Luciana Beatriz {\'A}vila, Am{\'a}lia Mendes, Iris Hendrickx

Challenges in modality annotation in a Brazilian Portuguese Spontaneous Speech Corpus

2013-03-01 · WS 2013 3 · Luciana Beatriz Avila, Heliana Mello
Sentiment Analysis

Modality in Text: a Proposal for Corpus Annotation

2012-05-01 · LREC 2012 5 · Iris Hendrickx, Am{\'a}lia Mendes, Silvia Mencarelli

We present a annotation scheme for modality in Portuguese. In our annotation scheme we have tried to combine a more theoretical linguistic viewpoint with a practical annotation scheme that will also be useful for NLP res…

Opinion MiningSentenceSentiment Analysis

Semantically Inspired AMR Alignment for the Portuguese Language

2020-11-01 · EMNLP 2020 11 · Rafael Anchi{\^e}ta, Thiago Pardo

Abstract Meaning Representation (AMR) is a graph-based semantic formalism where the nodes are concepts and edges are relations among them. Most of AMR parsing methods require alignment between the nodes of the graph and …

Abstract Meaning RepresentationAMR ParsingSentence

ACE-2005-PT: Corpus for Event Extraction in Portuguese

2024-08-29 · Luís Filipe Cunha, Purificação Silvano, Ricardo Campos, Alípio Jorge

Event extraction is an NLP task that commonly involves identifying the central word (trigger) for an event and its associated arguments in text. ACE-2005 is widely recognised as the standard corpus in this field. While o…

Event ExtractionLemmatization