paper-with-me

홈 › Papers

Pho(SC)-CTC -- A Hybrid Approach Towards Zero-shot Word Image Recognition

2021-05-31 · Ravi Bhatt, Anuj Rai, Narayanan C. Krishnan, Sukalpa Chanda

Annotating words in a historical document image archive for word image recognition purpose demands time and skilled human resource (like historians, paleographers). In a real-life scenario, obtaining sample images for all possible words is also not feasible. However, Zero-shot learning methods could aptly be used to recognize unseen/out-of-lexicon words in such historical document images. Based on previous state-of-the-art method for zero-shot word recognition Pho(SC)Net, we propose a hybrid model based on the CTC framework (Pho(SC)-CTC) that takes advantage of the rich features learned by Pho(SC)Net followed by a connectionist temporal classification (CTC) framework to perform the final classification. Encouraging results were obtained on two publicly available historical document datasets and one synthetic handwritten dataset, which justifies the efficacy of Pho(SC)-CTC and Pho(SC)Net.

📄 PDF Abstract BibTeX arXiv:2105.15093

Code (1)

anuj-rai-23/PHOSC-Zero-Shot-Word-Recognition 공식 구현 tf

Tasks

Zero-Shot Learning

Similar Papers 제목 키워드 기반

Alternative Semantic Representations for Zero-Shot Human Action Recognition

2017-06-28 · Qian Wang, Ke Chen

A proper semantic representation for encoding side information is key to the success of zero-shot learning. In this paper, we explore two alternative semantic representations especially for zero-shot human action recogni…

Action RecognitionTemporal Action LocalizationZero-Shot Action RecognitionZero-Shot Learning

A Fistful of Words: Learning Transferable Visual Models from Bag-of-Words Supervision

2021-12-27 · Ajinkya Tejankar, Maziar Sanjabi, Bichen Wu, Saining Xie 외

Using natural language as a supervision for training visual recognition models holds great promise. Recent works have shown that if such supervision is used in the form of alignment between images and captions in large t…

ClassificationImage Captioningimage-classificationImage Classification+3

Language Models as Zero-shot Visual Semantic Learners

2021-07-26 · Yue Jiao, Jonathon Hare, Adam Prügel-Bennett

Visual Semantic Embedding (VSE) models, which map images into a rich semantic embedding space, have been a milestone in object recognition and zero-shot learning. Current approaches to VSE heavily rely on static word em-…

ObjectObject RecognitionWord EmbeddingsZero-Shot Learning

TARN: Temporal Attentive Relation Network for Few-Shot and Zero-Shot Action Recognition

2019-07-21 · Mina Bishay, Georgios Zoumpourlis, Ioannis Patras

In this paper we propose a novel Temporal Attentive Relation Network (TARN) for the problems of few-shot and zero-shot action recognition. At the heart of our network is a meta-learning approach that learns to compare re…

Action RecognitionFew-Shot action recognitionFew Shot Action RecognitionMeta-Learning+4

Zero-Shot Learning Based Approach For Medieval Word Recognition Using Deep-Learned Features

2018-10-01 · 16th International Conference on Frontiers in Handwriting Recognition (ICFHR 2018) 2018 10 · Sukalpa Chanda, Jochem Baas, Daniël Haitink, Sebastien Hamely 외

Historical manuscripts reflect our past. Recently digitization of large quantities of historical handwritten docu- ments is taking place in every corner of the world, and are being archived. From those digital repositori…

AttributeGeneralized Zero-Shot LearningLanguage ModellingOptical Character Recognition (OCR)+2