paper-with-me

홈 › Papers

Making Metadata More FAIR Using Large Language Models

2023-07-24 · Sowmya S. Sundaram, Mark A. Musen

With the global increase in experimental data artifacts, harnessing them in a unified fashion leads to a major stumbling block - bad metadata. To bridge this gap, this work presents a Natural Language Processing (NLP) informed application, called FAIRMetaText, that compares metadata. Specifically, FAIRMetaText analyzes the natural language descriptions of metadata and provides a mathematical similarity measure between two terms. This measure can then be utilized for analyzing varied metadata, by suggesting terms for compliance or grouping similar terms for identification of replaceable terms. The efficacy of the algorithm is presented qualitatively and quantitatively on publicly available research artifacts and demonstrates large gains across metadata related tasks through an in-depth study of a wide variety of Large Language Models (LLMs). This software can drastically reduce the human effort in sifting through various natural language metadata while employing several experimental datasets on the same topic.

📄 PDF Abstract BibTeX arXiv:2307.13085

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

FAIR Metadata: A Community-driven Vocabulary Application

2021-11-06 · Christopher B. Rauch, Mat Kelly, John A. Kunze, Jane Greenberg

FAIR metadata is critical to supporting FAIR data overall. Transparency, community engagement, and flexibility are key aspects of FAIR that apply to metadata. This paper presents YAMZ (Yet Another Metadata Zoo), a commun…

Metadata Might Make Language Models Better

2022-11-18 · Kaspar Beelen, Daniel van Strien

This paper discusses the benefits of including metadata when training language models on historical collections. Using 19th-century newspapers as a case study, we extend the time-masking approach proposed by Rosin et al.…

Language ModelingLanguage Modelling

AutoFAIR : Automatic Data FAIRification via Machine Reading

2024-08-07 · Tingyan Ma, Wei Liu, Bin Lu, Xiaoying Gan 외

The explosive growth of data fuels data-driven research, facilitating progress across diverse domains. The FAIR principles emerge as a guiding standard, aiming to enhance the findability, accessibility, interoperability,…

FairnessReading Comprehension

Identity resolution of software metadata using Large Language Models

2025-05-29 · Eva Martín del Pico, Josep Lluís Gelpí, Salvador Capella-Gutiérrez

Software is an essential component of research. However, little attention has been paid to it compared with that paid to research data. Recently, there has been an increase in efforts to acknowledge and highlight the imp…

Fairness

Making Machine Learning Datasets and Models FAIR for HPC: A Methodology and Case Study

2022-11-03 · Pei-Hung Lin, Chunhua Liao, Winson Chen, Tristan Vanderbruggen 외

The FAIR Guiding Principles aim to improve the findability, accessibility, interoperability, and reusability of digital content by making them both human and machine actionable. However, these principles have not yet bee…

Fairness