paper-with-me

홈 › Papers

Enhancing Multimodal Entity and Relation Extraction with Variational Information Bottleneck

2023-04-05 · Shiyao Cui, Jiangxia Cao, Xin Cong, Jiawei Sheng, Quangang Li, Tingwen Liu, Jinqiao Shi

This paper studies the multimodal named entity recognition (MNER) and multimodal relation extraction (MRE), which are important for multimedia social platform analysis. The core of MNER and MRE lies in incorporating evident visual information to enhance textual semantics, where two issues inherently demand investigations. The first issue is modality-noise, where the task-irrelevant information in each modality may be noises misleading the task prediction. The second issue is modality-gap, where representations from different modalities are inconsistent, preventing from building the semantic alignment between the text and image. To address these issues, we propose a novel method for MNER and MRE by Multi-Modal representation learning with Information Bottleneck (MMIB). For the first issue, a refinement-regularizer probes the information-bottleneck principle to balance the predictive evidence and noisy information, yielding expressive representations for prediction. For the second issue, an alignment-regularizer is proposed, where a mutual information-based item works in a contrastive manner to regularize the consistent text-image representations. To our best knowledge, we are the first to explore variational IB estimation for MNER and MRE. Experiments show that MMIB achieves the state-of-the-art performances on three public benchmarks.

📄 PDF Abstract BibTeX arXiv:2304.02328

Code (0)

등록된 구현이 없습니다.

Tasks

named-entity-recognitionNamed Entity RecognitionRelationRelation ExtractionRepresentation Learning

Similar Papers 제목 키워드 기반

Prompt Me Up: Unleashing the Power of Alignments for Multimodal Entity and Relation Extraction

2023-10-25 · Xuming Hu, Junzhe Chen, Aiwei Liu, Shiao Meng 외

How can we better extract entities and relations from text? Using multimodal extraction with images and text obtains more signals for entities and relations, and aligns them through graphs or hierarchical fusion, aiding …

RelationRelation Extraction

Chain-of-Thought Prompt Distillation for Multimodal Named Entity Recognition and Multimodal Relation Extraction

2023-06-25 · Feng Chen, Yujian Feng

Multimodal Named Entity Recognition (MNER) and Multimodal Relation Extraction (MRE) necessitate the fundamental reasoning capacity for intricate linguistic and multimodal comprehension. In this study, we explore distilli…

Data AugmentationDomain Generalizationnamed-entity-recognitionNamed Entity Recognition+3

Variational Multi-Modal Hypergraph Attention Network for Multi-Modal Relation Extraction

2024-04-18 · Qian Li, Cheng Ji, Shu Guo, Yong Zhao 외

Multi-modal relation extraction (MMRE) is a challenging task that aims to identify relations between entities in text leveraging image information. Existing methods are limited by their neglect of the multiple entity pai…

DiversityRelationRelation ExtractionSentence

Multimodal Relational Triple Extraction with Query-based Entity Object Transformer

2024-08-16 · Lei Hei, Ning An, Tingjing Liao, Qi Ma 외

Multimodal Relation Extraction is crucial for constructing flexible and realistic knowledge graphs. Recent studies focus on extracting the relation type with entity pairs present in different modalities, such as one enti…

Knowledge GraphsObjectobject-detectionObject Detection+3

Joint Multimodal Entity-Relation Extraction Based on Edge-enhanced Graph Alignment Network and Word-pair Relation Tagging

2022-11-28 · Li Yuan, Yi Cai, Jin Wang, Qing Li

Multimodal named entity recognition (MNER) and multimodal relation extraction (MRE) are two fundamental subtasks in the multimodal knowledge graph construction task. However, the existing methods usually handle two tasks…

graph constructionnamed-entity-recognitionNamed Entity RecognitionNamed Entity Recognition (NER)+3