Key-value Pair Extraction
2개 벤치마크 · 논문 13편 · 이 태스크의 논문 보기 →
Benchmarks
Most implemented
LayoutXLM: Multimodal Pre-training for Multilingual Visually-rich Document Understanding
LiLT: A Simple yet Effective Language-Independent Layout Transformer for Structured Document Understanding
OCR-free Document Understanding Transformer
LayoutLMv3: Pre-training for Document AI with Unified Text and Image Masking
Papers
MedStruct-S: A Benchmark for Key Discovery, Key-Conditioned QA and Semi-Structured Extraction from OCR Clinical Reports
Semi-structured information extraction (IE) from OCR-derived clinical reports is crucial for efficiently reconstructing patients' longitudinal medical histories. In practice, this scenario commonly involves three tasks: …
Key-value Pair ExtractionInformation ExtractionQuestion AnsweringKVP10k : A Comprehensive Dataset for Key-Value Pair Extraction in Business Documents
In recent years, the challenge of extracting information from business documents has emerged as a critical task, finding applications across numerous domains. This effort has attracted substantial interest from both indu…
DiversityKey Information ExtractionKey-value Pair ExtractionUniVIE: A Unified Label Space Approach to Visual Information Extraction from Form-like Documents
Existing methods for Visual Information Extraction (VIE) from form-like documents typically fragment the process into separate subtasks, such as key information extraction, key-value pair extraction, and choice group ext…
DecoderFormKey Information ExtractionKey-value Pair Extraction+2PEneo: Unifying Line Extraction, Line Grouping, and Entity Linking for End-to-end Document Pair Extraction
Document pair extraction aims to identify key and value entities as well as their relationships from visually-rich documents. Most existing methods divide it into two separate tasks: semantic entity recognition (SER) and…
Key Information ExtractionKey-value Pair ExtractionRelation ExtractionSemantic entity labelingReading Order Matters: Information Extraction from Visually-rich Documents by Token Path Prediction
Recent advances in multimodal pre-trained models have significantly improved information extraction from visually-rich documents (VrDs), in which named entity recognition (NER) is treated as a sequence-labeling task of p…
Entity LinkingKey Information ExtractionKey-value Pair Extractionnamed-entity-recognition+9Towards Zero-shot Relation Extraction in Web Mining: A Multimodal Approach with Relative XML Path
The rapid growth of web pages and the increasing complexity of their structure poses a challenge for web mining models. Web mining models are required to understand the semi-structured web pages, particularly when little…
Contrastive LearningKey-value Pair ExtractionRelationRelation Extraction