Recognizing Multimodal Entailment
How information is created, shared and consumed has changed rapidly in recent decades, in part thanks to new social platforms and technologies on the web. With ever-larger amounts of unstructured and limited labels, organizing and reconciling information from different sources and modalities is a central challenge in machine learning. This cutting-edge tutorial aims to introduce the multimodal entailment task, which can be useful for detecting semantic alignments when a single modality alone does not suffice for a whole content understanding. Starting with a brief overview of natural language processing, computer vision, structured data and neural graph learning, we lay the foundations for the multimodal sections to follow. We then discuss recent multimodal learning literature covering visual, audio and language streams, and explore case studies focusing on tasks which require fine-grained understanding of visual and linguistic semantics question answering, veracity and hatred classification. Finally, we introduce a new dataset for recognizing multimodal entailment, exploring it in a hands-on collaborative section. Overall, this tutorial gives an overview of multimodal learning, introduces a multimodal entailment dataset, and encourages future research in the topic.
Code (0)
등록된 구현이 없습니다.
Tasks
Graph LearningQuestion AnsweringSimilar Papers 제목 키워드 기반
Visual Denotations for Recognizing Textual Entailment
In the logic approach to Recognizing Textual Entailment, identifying phrase-to-phrase semantic relations is still an unsolved problem. Resources such as the Paraphrase Database offer limited coverage despite their large …
Natural Language InferenceSemantic Composition蘊涵句型分析於改進中文文字蘊涵識別系統 (Entailment Analysis for Improving Chinese Recognizing Textual Entailment System) [In Chinese]
Recognizing Partial Textual Entailment
A Study of the Effect of Resolving Negation and Sentiment Analysis in Recognizing Text Entailment for Arabic
Recognizing the entailment relation showed that its influence to extract the semantic inferences in wide-ranging natural language processing domains (text summarization, question answering, etc.) and enhanced the results…
Natural Language InferenceNegationNegation DetectionQuestion Answering+3