paper-with-me

홈 › Papers

ActiveAnno: General-Purpose Document-Level Annotation Tool with Active Learning Integration

2021-06-01 · NAACL 2021 4 · Max Wiechmann, Seid Muhie Yimam, Chris Biemann

ActiveAnno is an annotation tool focused on document-level annotation tasks developed both for industry and research settings. It is designed to be a general-purpose tool with a wide variety of use cases. It features a modern and responsive web UI for creating annotation projects, conducting annotations, adjudicating disagreements, and analyzing annotation results. ActiveAnno embeds a highly configurable and interactive user interface. The tool also integrates a RESTful API that enables integration into other software systems, including an API for machine learning integration. ActiveAnno is built with extensible design and easy deployment in mind, all to enable users to perform annotation tasks with high efficiency and high-quality annotation results.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Active Learning

Similar Papers 제목 키워드 기반

MeDocVL: A Visual Language Model for Medical Document Understanding and Parsing

2026-02-06 · Wenjie Wang, Wei Wu, Ying Liu, Yuan Zhao 외 arxiv

Medical document OCR is challenging due to complex layouts, domain-specific terminology, and noisy annotations, while requiring strict field-level exact matching. Existing OCR systems and general-purpose vision-language …

Reinforcement Learning

Dr. DocBench: A Comprehensive Benchmark for Expert-Level and Difficult Document Parsing

2026-05-31 · Minglai Yang, Xinyan Velocity Yu, Pengyuan Li, Xinyu Guo 외 arxiv

Document parsing and recognition are fundamental capabilities for vision-language models (VLMs) and document processing systems. However, existing Optical Character Recognition (OCR) and document parsing benchmarks are i…

docExtractor: An off-the-shelf historical document element extraction

2020-12-15 · Tom Monnier, Mathieu Aubry

We present docExtractor, a generic approach for extracting visual elements such as text lines or illustrations from historical documents without requiring any real data annotation. We demonstrate it provides high-quality…

Document Layout AnalysisSegmentation

propella-1: Multi-Property Document Annotation for LLM Data Curation at Scale

2026-02-12 · Maximilian Idahl, Benedikt Droste, Björn Plüster, Jan Philipp Harries arxiv

Since FineWeb-Edu, data curation for LLM pretraining has predominantly relied on single scalar quality scores produced by small classifiers. A single score conflates multiple quality dimensions, prevents flexible filteri…

Annotation Study of Japanese Judgments on Tort for Legal Judgment Prediction with Rationales

2022-06-01 · LREC 2022 6 · Hiroaki Yamada, Takenobu Tokunaga, Ryutaro Ohara, Keisuke Takeshita 외

This paper describes a comprehensive annotation study on Japanese judgment documents in civil cases. We aim to build an annotated corpus designed for Legal Judgment Prediction (LJP), especially for torts. Our annotation …