paper-with-me

Papers

Visually Guided Spatial Relation Extraction from Text

2018-06-01 · NAACL 2018 6 · Taher Rahgooy, Umar Manzoor, Parisa Kordjamshidi

Extraction of spatial relations from sentences with complex/nesting relationships is very challenging as often needs resolving inherent semantic ambiguities. We seek help from visual modality to fill the information gap in the text modality and resolve spatial semantic ambiguities. We use various recent vision and language datasets and techniques to train inter-modality alignment models, visual relationship classifiers and propose a novel global inference model to integrate these components into our structured output prediction model for spatial role and relation extraction. Our global inference model enables us to utilize the visual and geometric relationships between objects and improves the state-of-art results of spatial information extraction from text.

📄 PDF Abstract BibTeX

Code (0)

등록된 구현이 없습니다.

Tasks

Activity RecognitionImage CaptioningImage RetrievalObject LocalizationQuestion AnsweringRelationRelation ExtractionVisual Question Answering (VQA)

Similar Papers 제목 키워드 기반

RE$^2$: Region-Aware Relation Extraction from Visually Rich Documents

2023-05-24 · Pritika Ramu, Sijia Wang, Lalla Mouatadid, Joy Rimchala 외

Current research in form understanding predominantly relies on large pre-trained language models, necessitating extensive data for pre-training. However, the importance of layout structure (i.e., the spatial relationship…

Graph AttentionRelationRelation ExtractionRelation Prediction

Multimodal weighted graph representation for information extraction from visually rich documents.

2024-01-05 · Neurocomputing 2024 1 · Hamza Gbada, Karim Kalti, Mohamed Ali Mahjoub

This paper introduces a novel system for information extraction from visually rich documents (VRD) using a weighted graph representation. The proposed method aims to improve the performance of the information extraction …

Document Layout Analysisdocument understandingGraph Neural NetworkInformation Retrieval+2

Towards Human-Like Machine Comprehension: Few-Shot Relational Learning in Visually-Rich Documents

2024-03-23 · Hao Wang, Tang Li, Chenhui Chu, Nengjun Zhu 외

Key-value relations are prevalent in Visually-Rich Documents (VRDs), often depicted in distinct spatial regions accompanied by specific color and font styles. These non-textual cues serve as important indicators that gre…

Document AIReading ComprehensionRelationRelational Reasoning

A LayoutLMv3-Based Model for Enhanced Relation Extraction in Visually-Rich Documents

2024-04-16 · Wiam Adnan, Joel Tang, Yassine Bel Khayat Zouggari, Seif Edinne Laatiri 외

Document Understanding is an evolving field in Natural Language Processing (NLP). In particular, visual and spatial features are essential in addition to the raw text itself and hence, several multimodal models were deve…

document understandingKey Information ExtractionRelationRelation Extraction

Global Structure Knowledge-Guided Relation Extraction Method for Visually-Rich Document

2023-05-23 · Xiangnan Chen, Qian Xiao, Juncheng Li, Duo Dong 외

Visual Relation Extraction (VRE) is a powerful means of discovering relationships between entities within visually-rich documents. Existing methods often focus on manipulating entity features to find pairwise relations, …

RelationRelation Extraction