paper-with-me

홈 › Papers

YEDDA: A Lightweight Collaborative Text Span Annotation Tool

2017-11-10 · ACL 2018 7 · Jie Yang, Yue Zhang, Linwei Li, Xingxuan Li

In this paper, we introduce \textsc{Yedda}, a lightweight but efficient and comprehensive open-source tool for text span annotation. \textsc{Yedda} provides a systematic solution for text span annotation, ranging from collaborative user annotation to administrator evaluation and analysis. It overcomes the low efficiency of traditional text annotation tools by annotating entities through both command line and shortcut keys, which are configurable with custom labels. \textsc{Yedda} also gives intelligent recommendations by learning the up-to-date annotated text. An administrator client is developed to evaluate annotation quality of multiple annotators and generate detailed comparison report for each annotator pair. Experiments show that the proposed system can reduce the annotation time by half compared with existing annotation tools. And the annotation time can be further compressed by 16.47\% through intelligent recommendation.

📄 PDF Abstract BibTeX arXiv:1711.03759

Code (1)

jiesutd/YEDDA 공식 구현

Tasks

text annotation

Similar Papers 제목 키워드 기반

SLATE: A Super-Lightweight Annotation Tool for Experts

2019-07-18 · ACL 2019 7 · Jonathan K. Kummerfeld

Many annotation tools have been developed, covering a wide variety of tasks and providing features like user management, pre-processing, and automatic labeling. However, all of these tools use Graphical User Interfaces, …

Management

Chinese-SkillSpan: A Span-Level Dataset for ESCO-Aligned Competency Extraction from Chinese Job Ads

2026-04-24 · Guojing Li, Zichuan Fu, Junyi Li, Wenxia Zhou 외 arxiv

Job Skill Named Entity Recognition (JobSkillNER) aims to automatically extract key skill information from large-scale job posting data, which is important for improving talent-market matching efficiency and supporting pe…

Scalable and Culturally Specific Stereotype Dataset Construction via Human-LLM Collaboration

2026-07-08 · Weicheng Ma, John Guerrerio, Soroush Vosoughi arxiv

Research on stereotypes in large language models (LLMs) has largely focused on English-speaking contexts, due to the lack of datasets in other languages and the high cost of manual annotation in underrepresented cultures…

From Documents to Spans: Scalable Supervision for Evidence-Based ICD Coding with LLMs

2026-03-16 · Xu Zhang, Wenxin Ma, Chenxu Wu, Rongsheng Wang 외 arxiv

International Classification of Diseases (ICD) coding assigns diagnosis codes to clinical documents and is essential for healthcare billing and clinical analysis. Reliable coding requires that each predicted code be supp…

ACAT: A Collaborative Platform for Efficient Aspect-Based Sentiment Dataset Annotation

2026-06-02 · Ana-Maria Luisa Mocanu, Ciprian-Octavian Truica, Elena-Simona Apostol arxiv

Aspect-Based Sentiment Analysis (ABSA) requires high-quality datasets to train reliable models. However, existing annotation tools treat output as flat files, leaving researchers to manually consolidate multi-annotator d…

Aspect Sentiment Triplet ExtractionSentiment Analysis