FrSemCor: Annotating a French Corpus with Supersenses
French, as many languages, lacks semantically annotated corpus data. Our aim is to provide the linguistic and NLP research communities with a gold standard sense-annotated corpus of French, using WordNet Unique Beginners as semantic tags, thus allowing for interoperability. In this paper, we report on the first phase of the project, which focused on the annotation of common nouns. The resulting dataset consists of more than 12,000 French noun occurrences which were annotated in double blind and adjudicated according to a carefully redefined set of supersenses. The resource is released online under a Creative Commons Licence.
Code (0)
등록된 구현이 없습니다.
Similar Papers 제목 키워드 기반
PASTRIE: A Corpus of Prepositions Annotated with Supersense Tags in Reddit International English
We present the Prepositions Annotated with Supersense Tags in Reddit International English ("PASTRIE") corpus, a new dataset containing manually annotated preposition supersenses of English data from presumed speakers of…
Cr\'eation d'un multi-arbre \`a partir d'un texte balis\'e : l'exemple de l'annotation d'un corpus d'oral spontan\'e (Creating a Multi-Tree from a Tagged Text : Annotating Spoken French) [in French]
A corpus of preposition supersenses in English web reviews
We present the first corpus annotated with preposition supersenses, unlexicalized categories for semantic functions that can be marked by English prepositions (Schneider et al., 2015). That scheme improves upon its prede…
Adpositional Supersenses for Mandarin Chinese
This study adapts Semantic Network of Adposition and Case Supersenses (SNACS) annotation to Mandarin Chinese and demonstrates that the same supersense categories are appropriate for Chinese adposition semantics. We annot…
Machine TranslationTranslationding-01 :ARG0: An AMR Corpus for Spontaneous French Dialogue
We present our work to build a French semantic corpus by annotating French dialogue in Abstract Meaning Representation (AMR). Specifically, we annotate the DinG corpus, consisting of transcripts of spontaneous French dia…