paper-with-me

홈 › Papers

CAPTION: Correction by Analyses, POS-Tagging and Interpretation of Objects using only Nouns

2020-10-02 · Leonardo Anjoletto Ferreira, Douglas De Rizzo Meneghetti, Paulo Eduardo Santos

Recently, Deep Learning (DL) methods have shown an excellent performance in image captioning and visual question answering. However, despite their performance, DL methods do not learn the semantics of the words that are being used to describe a scene, making it difficult to spot incorrect words used in captions or to interchange words that have similar meanings. This work proposes a combination of DL methods for object detection and natural language processing to validate image's captions. We test our method in the FOIL-COCO data set, since it provides correct and incorrect captions for various images using only objects represented in the MS-COCO image data set. Results show that our method has a good overall performance, in some cases similar to the human performance.

📄 PDF Abstract BibTeX arXiv:2010.00839

Code (0)

등록된 구현이 없습니다.

Tasks

Image Captioningobject-detectionObject DetectionPOSPOS TaggingQuestion AnsweringVisual Question AnsweringVisual Question Answering (VQA)

Similar Papers 제목 키워드 기반

View Selection for 3D Captioning via Diffusion Ranking

2024-04-11 · Tiange Luo, Justin Johnson, Honglak Lee

Scalable annotation approaches are crucial for constructing extensive 3D-text datasets, facilitating a broader range of applications. However, existing methods sometimes lead to the generation of hallucinated captions, c…

3D Object CaptioningHallucinationImage CaptioningQuestion Answering+2

ProCap: Prominence-guided Object Rectification for Faithful and Comprehensive Video Captioning

2026-07-23 · Debjyoti Das Adhikary, Aritra Hazra, Partha Pratim Chakrabarti arxiv

Improving video captioning quality typically demands retraining large vision-language models, an expensive and often impractical requirement. Existing training-free alternatives instead ground captions in detected object…

Video Captioning

BERT Enhanced Neural Machine Translation and Sequence Tagging Model for Chinese Grammatical Error Diagnosis

2020-12-01 · AACL (NLP-TEA) 2020 12 · Deng Liang, Chen Zheng, Lei Guo, Xin Cui 외

This paper presents the UNIPUS-Flaubert team’s hybrid system for the NLPTEA 2020 shared task of Chinese Grammatical Error Diagnosis (CGED). As a challenging NLP task, CGED has attracted increasing attention recently and …

Grammatical Error CorrectionMachine TranslationNMTTranslation

Correcting Errors in a New Gold Standard for Tagging Icelandic Text

2014-05-01 · LREC 2014 5 · Sigr{\'u}n Helgad{\'o}ttir, Hrafn Loftsson, Eir{\'\i}kur R{\"o}gnvaldsson

In this paper, we describe the correction of PoS tags in a new Icelandic corpus, MIM-GOLD, consisting of about 1 million tokens sampled from the Tagged Icelandic Corpus, M{\'I}M, released in 2013. The goal is to use the …

Part-Of-Speech TaggingPOS

Music autotagging as captioning

2020-10-01 · NLP4MusA 2020 10 · Tian Cai, Michael I Mandel, Di He