paper-with-me

홈 › Papers

CTAP for Italian: Integrating Components for the Analysis of Italian into a Multilingual Linguistic Complexity Analysis Tool

2020-05-01 · LREC 2020 5 · Nadezda Okinina, Jennifer-Carmen Frey, Zarah Weiss

Linguistic complexity research being a very actively developing field, an increasing number of text analysis tools are created that use natural language processing techniques for the automatic extraction of quantifiable measures of linguistic complexity. While most tools are designed to analyse only one language, the CTAP open source linguistic complexity measurement tool is capable of processing multiple languages, making cross-lingual comparisons possible. Although it was originally developed for English, the architecture has been ex-tended to support multi-lingual analyses. Here we present the Italian component of CTAP, describe its implementation and compare it to the existing linguistic complexity tools for Italian. Offering general text length statistics and features for lexical, syntactic, and morpho-syntactic complexity (including measures of lexical frequency, lexical diversity, lexical and syntactical variation, part-of-speech density), CTAP is currently the most comprehensive linguistic complexity measurement tool for Italian and the only one allowing the comparison of Italian texts to multiple other languages within one tool.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Diversity

Similar Papers 제목 키워드 기반

Dialogue Act and Slot Recognition in Italian Complex Dialogues

2022-06-01 · EURALI (LREC) 2022 6 · Irene Sucameli, Michele De Quattro, Arash Eshghi, Alessandro Suglia 외

Since the advent of Transformer-based, pretrained language models (LM) such as BERT, Natural Language Understanding (NLU) components in the form of Dialogue Act Recognition (DAR) and Slot Recognition (SR) for dialogue sy…

Natural Language Understanding

Prior Polarity Lexical Resources for the Italian Language

2015-07-01 · Valeria Borzì, Simone Faro, Arianna Pavone, Sabrina Sansone

In this paper we present SABRINA (Sentiment Analysis: a Broad Resource for Italian Natural language Applications) a manually annotated prior polarity lexical resource for Italian natural language applications in the fiel…

Opinion MiningSentiment Analysis

Italian VerbNet: A Construction-based Approach to Italian Verb Classification

2016-05-01 · LREC 2016 5 · Lucia Busso, Aless Lenci, ro

This paper proposes a new method for Italian verb classification -and a preliminary example of resulting classes- inspired by Levin (1993) and VerbNet (Kipper-Schuler, 2005), yet partially independent from these resource…

ClassificationGeneral Classification

The ProLiFIC dataset: Leveraging LLMs to Unveil the Italian Lawmaking Process

2025-08-25 · Matilde Contestabile, Chiara Ferrara, Alberto Giovannetti, Giovanni Parrillo 외 arxiv

Process Mining (PM), initially developed for industrial and business contexts, has recently been applied to social systems, including legal ones. However, PM's efficacy in the legal domain is limited by the accessibility…

FEEL-IT: Emotion and Sentiment Classification for the Italian Language

2021-04-01 · EACL (WASSA) 2021 4 · Federico Bianchi, Debora Nozza, Dirk Hovy

While sentiment analysis is a popular task to understand people’s reactions online, we often need more nuanced information: is the post negative because the user is angry or sad? An abundance of approaches have been intr…

ClassificationSentiment AnalysisSentiment Classification