paper-with-me

홈 › Papers

E:Calm Resource: a Resource for Studying Texts Produced by French Pupils and Students

2020-05-01 · LREC 2020 5 · Lydia-Mai Ho-Dac, Serge Fleury, Claude Ponton

The E:Calm resource is constructed from French student texts produced in a variety of usual contexts of teaching. The distinction of the E:Calm resource is to provide an ecological data set that gives a broad overview of texts written at elementary school, high school and university. This paper describes the whole data processing: encoding of the main graphical aspects of the handwritten primary sources according to the TEI-P5 norm; spelling standardizing; POS tagging and syntactic parsing evaluation.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

POSPOS Tagging

Similar Papers 제목 키워드 기반

Extracting Linguistic Resources from the Web for Concept-to-Text Generation

2018-10-31 · Gerasimos Lampouras, Ion Androutsopoulos

Many concept-to-text generation systems require domain-specific linguistic resources to produce high quality texts, but manually constructing these resources can be tedious and costly. Focusing on NaturalOWL, a publicly …

Concept-To-Text GenerationSentenceText Generation

Theoretical and Methodological Framework for Studying Texts Produced by Large Language Models

2024-08-29 · Jiří Milička

This paper addresses the conceptual, methodological and technical challenges in studying large language models (LLMs) and the texts they produce from a quantitative linguistics perspective. It builds on a theoretical fra…

New language resources for the Pashto language

2012-05-01 · LREC 2012 5 · Djamel Mostefa, Khalid Choukri, Sylvie Brunessaux, Karim Boudahmane

This paper reports on the development of new language resources for the Pashto language, a very low-resource language spoken in Afghanistan and Pakistan. In the scope of a multilingual data collection project, three larg…

Machine TranslationTranslation

DimStance: Multilingual Datasets for Dimensional Stance Analysis

2026-01-29 · Jonas Becker, Liang-Chih Yu, Shamsuddeen Hassan Muhammad, Jan Philip Wahle 외 arxiv

Stance detection is an established task that classifies an author's attitude toward a specific target into categories such as Favor, Neutral, and Against. Beyond categorical stance labels, we leverage a long-established …

Stance Detection

ChiSCor: A Corpus of Freely Told Fantasy Stories by Dutch Children for Computational Linguistics and Cognitive Science

2023-10-31 · Bram M. A. van Dijk, Max J. van Duijn, Suzan Verberne, Marco R. Spruit

In this resource paper we release ChiSCor, a new corpus containing 619 fantasy stories, told freely by 442 Dutch children aged 4-12. ChiSCor was compiled for studying how children render character perspectives, and unrav…

LEMMAvalid