paper-with-me

Papers

Learning Deep Representations for Word Spotting Under Weak Supervision

2017-12-01 · Neha Gurjar, Sebastian Sudholt, Gernot A. Fink

Convolutional Neural Networks have made their mark in various fields of computer vision in recent years. They have achieved state-of-the-art performance in the field of document analysis as well. However, CNNs require a large amount of annotated training data and, hence, great manual effort. In our approach, we introduce a method to drastically reduce the manual annotation effort while retaining the high performance of a CNN for word spotting in handwritten documents. The model is learned with weak supervision using a combination of synthetically generated training data and a small subset of the training partition of the handwritten data set. We show that the network achieves results highly competitive to the state-of-the-art in word spotting with shorter training times and a fraction of the annotation effort.

📄 PDF Abstract BibTeX arXiv:1712.00250

Code (0)

등록된 구현이 없습니다.

Tasks

Word Spotting In Handwritten Documents

Similar Papers 제목 키워드 기반

Improving Small Footprint Few-shot Keyword Spotting with Supervision on Auxiliary Data

2023-08-31 · Seunghan Yang, Byeonggeun Kim, Kyuhong Shim, Simyung Chang

Few-shot keyword spotting (FS-KWS) models usually require large-scale annotated datasets to generalize to unseen target keywords. However, existing KWS datasets are limited in scale and gathering keyword-like labeled dat…

Keyword SpottingMulti-Task LearningSelf-Supervised Learning

Scaling up sign spotting through sign language dictionaries

2022-05-09 · Gül Varol, Liliane Momeni, Samuel Albanie, Triantafyllos Afouras 외

The focus of this work is $\textit{sign spotting}$ - given a video of an isolated sign, our task is to identify $\textit{whether}$ and $\textit{where}$ it has been signed in a continuous, co-articulated sign language vid…

Multiple Instance Learning

Watch, read and lookup: learning to spot signs from multiple supervisors

2020-10-08 · Liliane Momeni, Gül Varol, Samuel Albanie, Triantafyllos Afouras 외

The focus of this work is sign spotting - given a video of an isolated sign, our task is to identify whether and where it has been signed in a continuous, co-articulated sign language video. To achieve this sign spotting…

Multiple Instance Learning

SMART: MLLM-guided Temporal Alignment for Unifying Sign Language Recognition and Spotting

2026-08-26 · Eunjee Choi, JungHoon Sung, Seongwhan Cho, Chu Xin 외 arxiv

Continuous sign language recognition (CSLR) aims to recognize gloss sequences from unsegmented sign videos under weak sequence-level supervision. However, existing methods rely on sentence-level gloss annotations, provid…

Sign Language RecognitionRepresentation Learning

Multi-task Voice Activated Framework using Self-supervised Learning

2021-10-03 · Shehzeen Hussain, Van Nguyen, Shuhua Zhang, Erik Visser

Self-supervised learning methods such as wav2vec 2.0 have shown promising results in learning speech representations from unlabelled and untranscribed speech data that are useful for speech recognition. Since these repre…

Emotion ClassificationKeyword SpottingMulti-Task LearningSelf-Supervised Learning+3