paper-with-me

홈 › Papers

Quevedo: Annotation and Processing of Graphical Languages

2022-06-01 · LREC 2022 6 · Antonio F. G. Sevilla, Alberto Díaz Esteban, José María Lahoz-Bengoechea

In this article, we present Quevedo, a software tool we have developed for the task of automatic processing of graphical languages. These are languages which use images to convey meaning, relying not only on the shape of symbols but also on their spatial arrangement in the page, and relative to each other. When presented in image form, these languages require specialized computational processing which is not the same as usually done either for natural language processing or for artificial vision. Quevedo enables this specialized processing, focusing on a data-based approach. As a command line application and library, it provides features for the collection and management of image datasets, and their machine learning recognition using neural networks and recognizer pipelines. This processing requires careful annotation of the source data, for which Quevedo offers an extensive and visual web-based annotation interface. In this article, we also briefly present a case study centered on the task of SignWriting recognition, the original motivation for writing the software. Quevedo is written in Python, and distributed freely under the Open Software License version 3.0.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Management

Similar Papers 제목 키워드 기반

TextPro-AL: An Active Learning Platform for Flexible and Efficient Production of Training Data for NLP Tasks

2016-12-01 · COLING 2016 12 · Bernardo Magnini, Anne-Lyse Minard, Mohammed R. H. Qwaider, Manuela Speranza

This paper presents TextPro-AL (Active Learning for Text Processing), a platform where human annotators can efficiently work to produce high quality training data for new domains and new languages exploiting Active Learn…

Active LearningBIG-bench Machine LearningDomain Adaptation

Neural Graphical Models over Strings for Principal Parts Morphological Paradigm Completion

2017-04-01 · EACL 2017 4 · Ryan Cotterell, John Sylak-Glassman, Christo Kirov

Many of the world{'}s languages contain an abundance of inflected forms for each lexeme. A critical task in processing such languages is predicting these inflected forms. We develop a novel statistical model for the prob…

Morphological Analysis

CCGweb: a New Annotation Tool and a First Quadrilingual CCG Treebank

2019-08-01 · WS 2019 8 · Kilian Evang, Lasha Abzianidze, Johan Bos

We present the first open-source graphical annotation tool for combinatory categorial grammar (CCG), and the first set of detailed guidelines for syntactic annotation with CCG, for four languages: English, German, Italia…

Crossmodal-3600: A Massively Multilingual Multimodal Evaluation Dataset

2022-05-25 · Ashish V. Thapliyal, Jordi Pont-Tuset, Xi Chen, Radu Soricut

Research in massively multilingual image captioning has been severely hampered by a lack of high-quality evaluation datasets. In this paper we present the Crossmodal-3600 dataset (XM3600 in short), a geographically diver…

Image CaptioningImage RetrievalImage-text RetrievalImage-to-Text Retrieval+3

Deploying Technology to Save Endangered Languages

2019-08-23 · Hilaria Cruz, Joseph Waring

Computer scientists working on natural language processing, native speakers of endangered languages, and field linguists to discuss ways to harness Automatic Speech Recognition, especially neural networks, to automate an…

Automatic Speech RecognitionAutomatic Speech Recognition (ASR)speech-recognitionSpeech Recognition