paper-with-me

Papers

The Multilingual Anonymisation Toolkit for Public Administrations (MAPA) Project

2020-11-01 · EAMT 2020 11 · Ēriks Ajausks, Victoria Arranz, Laurent Bié, Aleix Cerdà-i-Cucó, Khalid Choukri, Montse Cuadros, Hans Degroote, Amando Estela, Thierry Etchegoyhen, Mercedes García-Martínez, Aitor García-Pablos, Manuel Herranz, Alejandro Kohan, Maite Melero, Mike Rosner, Roberts Rozis, Patrick Paroubek, Artūrs Vasiļevskis, Pierre Zweigenbaum

We describe the MAPA project, funded under the Connecting Europe Facility programme, whose goal is the development of an open-source de-identification toolkit for all official European Union languages. It will be developed since January 2020 until December 2021.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

De-identification

Similar Papers 제목 키워드 기반

Spanish Datasets for Sensitive Entity Detection in the Legal Domain

2022-06-01 · LREC 2022 6 · Ona de Gibert Bonet, Aitor García Pablos, Montse Cuadros, Maite Melero

The de-identification of sensible data, also known as automatic textual anonymisation, is essential for data sharing and reuse, both for research and commercial purposes. The first step for data anonymisation is the dete…

De-identification

Naamapadam: A Large-Scale Named Entity Annotated Data for Indic Languages

2022-12-20 · Arnav Mhaske, Harshit Kedia, Sumanth Doddapaneni, Mitesh M. Khapra 외

We present, Naamapadam, the largest publicly available Named Entity Recognition (NER) dataset for the 11 major Indian languages from two language families. The dataset contains more than 400k sentences annotated with a t…

Named Entity RecognitionNamed Entity Recognition (NER)Sentence

MAPA Project: Ready-to-Go Open-Source Datasets and Deep Learning Technology to Remove Identifying Information from Text Documents

2022-06-01 · LEGAL (LREC) 2022 6 · Victoria Arranz, Khalid Choukri, Montse Cuadros, Aitor García Pablos 외

This paper presents the outcomes of the MAPA project, a set of annotated corpora for 24 languages of the European Union and an open-source customisable toolkit able to detect and substitute sensitive information in text …

De-identificationnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)

MapAnything: Evaluating Monocular Metric Depth Models for 3D Urban Asset Localization

2025-09-18 · Miriam Louise Carnot, Jonas Kunze, Erik Quinten Fastermann, Eric Peukert 외 arxiv

City administrations increasingly rely on comprehensive databases and digital twins of city assets, such as traffic signs and trees, as well as incidents such as graffiti or road damage, to maintain an effective overview…

Depth EstimationPoint Clouds

Automated Anonymisation of Visual and Audio Data in Classroom Studies

2020-01-14 · Ömer Sümer, Peter Gerjets, Ulrich Trautwein, Enkelejda Kasneci

Understanding students' and teachers' verbal and non-verbal behaviours during instruction may help infer valuable information regarding the quality of teaching. In education research, there have been many studies that ai…

Management