AutoFAIR : Automatic Data FAIRification via Machine Reading
The explosive growth of data fuels data-driven research, facilitating progress across diverse domains. The FAIR principles emerge as a guiding standard, aiming to enhance the findability, accessibility, interoperability, and reusability of data. However, current efforts primarily focus on manual data FAIRification, which can only handle targeted data and lack efficiency. To address this issue, we propose AutoFAIR, an architecture designed to enhance data FAIRness automately. Firstly, We align each data and metadata operation with specific FAIR indicators to guide machine-executable actions. Then, We utilize Web Reader to automatically extract metadata based on language models, even in the absence of structured data webpage schemas. Subsequently, FAIR Alignment is employed to make metadata comply with FAIR principles by ontology guidance and semantic matching. Finally, by applying AutoFAIR to various data, especially in the field of mountain hazards, we observe significant improvements in findability, accessibility, interoperability, and reusability of data. The FAIRness scores before and after applying AutoFAIR indicate enhanced data value.
Code (0)
등록된 구현이 없습니다.
Tasks
FairnessReading ComprehensionMethods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
A Chinese Machine Reading Comprehension Dataset Automatic Generated Based on Knowledge Graph
“Machine reading comprehension (MRC) is a typical natural language processing (NLP)task and has developed rapidly in the last few years. Various reading comprehension datasets have been built to support MRC studies. Howe…
Dataset GenerationMachine Reading ComprehensionReading ComprehensionKorQuAD1.0: Korean QA Dataset for Machine Reading Comprehension
Machine Reading Comprehension (MRC) is a task that requires machine to understand natural language and answer questions by reading a document. It is the core of automatic response technology such as chatbots and automati…
ArticlesMachine Reading ComprehensionQuestion AnsweringReading ComprehensionQuantifying sentence complexity based on eye-tracking measures
Eye-tracking reading times have been attested to reflect cognitive processes underlying sentence comprehension. However, the use of reading times in NLP applications is an underexplored area of research. In this initial …
Part-Of-Speech TaggingSarcasm DetectionSentenceText SimplificationMachine Reading Tea Leaves: Automatically Evaluating Topic Coherence and Topic Model Quality
Music Proofreading with RefinPaint: Where and How to Modify Compositions given Context
Autoregressive generative transformers are key in music generation, producing coherent compositions but facing challenges in human-machine collaboration. We propose RefinPaint, an iterative technique that improves the sa…
Music Generation