paper-with-me

홈 › Papers

Experimenting active and sequential learning in a medieval music manuscript

2025-07-21 · Sachin Sharma, Federico Simonetta, Michele Flammini arxiv

Optical Music Recognition (OMR) is a cornerstone of music digitization initiatives in cultural heritage, yet it remains limited by the scarcity of annotated data and the complexity of historical manuscripts. In this paper, we present a preliminary study of Active Learning (AL) and Sequential Learning (SL) tailored for object detection and layout recognition in an old medieval music manuscript. Leveraging YOLOv8, our system selects samples with the highest uncertainty (lowest prediction confidence) for iterative labeling and retraining. Our approach starts with a single annotated image and successfully boosts performance while minimizing manual labeling. Experimental results indicate that comparable accuracy to fully supervised training can be achieved with significantly fewer labeled examples. We test the methodology as a preliminary investigation on a novel dataset offered to the community by the Anonymous project, which studies laude, a poetical-musical genre spread across Italy during the 12th-16th Century. We show that in the manuscript at-hand, uncertainty-based AL is not effective and advocates for more usable methods in data-scarcity scenarios.

📄 PDF Abstract BibTeX arXiv:2507.15633

Code (0)

등록된 구현이 없습니다.

Tasks

Object DetectionActive Learning

Similar Papers 제목 키워드 기반

Exploring the "Great Unseen" in Medieval Manuscripts: Instance-Level Labeling of Legacy Image Collections with Zero-Shot Models

2025-11-10 · Christofer Meinecke, Estelle Guéville, David Joseph Wrisley arxiv

We aim to theorize the medieval manuscript page and its contents more holistically, using state-of-the-art techniques to segment and describe the entire manuscript folio, for the purpose of creating richer training data …

Instance Segmentation

Co-Occurrence Patterns in the Voynich Manuscript

2016-01-27 · Torsten Timm

The Voynich Manuscript is a medieval book written in an unknown script. This paper studies the distribution of similarly spelled words in the Voynich Manuscript. It shows that the distribution of words within the manuscr…

DIVA-HisDB: A Precisely Annotated Large Dataset of Challenging Medieval Manuscripts

2016-10-23 · International Conference on Frontiers in Handwriting Recognition 2016 10 · Fotini Simistira, Mathias Seuret, Nicole Eichenberger, Angelika Garz 외

This paper introduces a publicly available historical manuscript database DIVA-HisDB for the evaluation of several Document Image Analysis (DIA) tasks. The database consists of 150 annotated pages of three different medi…

BinarizationDocument Layout AnalysisSegmentationText-Line Extraction

How the Voynich Manuscript was created

2014-07-24 · Torsten Timm

The Voynich manuscript is a medieval book written in an unknown script. This paper studies the relation between similarly spelled words in the Voynich manuscript. By means of a detailed analysis of similar spelled words …

RelationText Generation

Enriching Digitized Medieval Manuscripts: Linking Image, Text and Lexical Knowledge

2015-07-01 · WS 2015 7 · Aitor Arronte {\'A}lvarez