paper-with-me

Papers

Data Overdose? Time for a Quadruple Shot: Knowledge Graph Construction using Enhanced Triple Extraction

2025-08-05 · Taine J. Elliott, Stephen P. Levitt, Ken Nixon, Martin Bekker arxiv

The rapid expansion of publicly-available medical data presents a challenge for clinicians and researchers alike, increasing the gap between the volume of scientific literature and its applications. The steady growth of studies and findings overwhelms medical professionals at large, hindering their ability to systematically review and understand the latest knowledge. This paper presents an approach to information extraction and automatic knowledge graph (KG) generation to identify and connect biomedical knowledge. Through a pipeline of large language model (LLM) agents, the system decomposes 44 PubMed abstracts into semantically meaningful proposition sentences and extracts KG triples from these sentences. The triples are enhanced using a combination of open domain and ontology-based information extraction methodologies to incorporate ontological categories. On top of this, a context variable is included during extraction to allow the triple to stand on its own - thereby becoming `quadruples'. The extraction accuracy of the LLM is validated by comparing natural language sentences generated from the enhanced triples to the original propositions, achieving an average cosine similarity of 0.874. The similarity for generated sentences of enhanced triples were compared with generated sentences of ordinary triples showing an increase as a result of the context variable. Furthermore, this research explores the ability for LLMs to infer new relationships and connect clusters in the knowledge base of the knowledge graph. This approach leads the way to provide medical practitioners with a centralised, updated in real-time, and sustainable knowledge source, and may be the foundation of similar gains in a wide variety of fields.

📄 PDF Abstract BibTeX arXiv:2508.03438

Code (0)

등록된 구현이 없습니다.

Tasks

Information Extraction

Similar Papers 제목 키워드 기반

Large Language Models for Drug Overdose Prediction from Longitudinal Medical Records

2025-04-16 · Md Sultan Al Nahian, Chris Delcher, Daniel Harris, Peter Akpunonu 외

The ability to predict drug overdose risk from a patient's medical records is crucial for timely intervention and prevention. Traditional machine learning models have shown promise in analyzing longitudinal medical recor…

Point Process Modeling of Drug Overdoses with Heterogeneous and Missing Data

2020-10-12 · Xueying Liu, Jeremy Carter, Brad Ray, George Mohler

Opioid overdose rates have increased in the United States over the past decade and reflect a major public health crisis. Modeling and prediction of drug and opioid hotspots, where a high percentage of events fall in a sm…

ClusteringPoint Processes

The Limits of ChatGPT in Extracting Aspect-Category-Opinion-Sentiment Quadruples: A Comparative Analysis

2023-10-10 · Xiancai Xu, Jia-Dong Zhang, Rongchang Xiao, Lei Xiong

Recently, ChatGPT has attracted great attention from both industry and academia due to its surprising abilities in natural language understanding and generation. We are particularly curious about whether it can achieve p…

Aspect-Based Sentiment AnalysisIn-Context LearningNatural Language UnderstandingSentiment Analysis

Machine Learning for Drug Overdose Surveillance

2017-10-06 · Daniel B. Neill, William Herlands

We describe two recently proposed machine learning approaches for discovering emerging trends in fatal accidental drug overdoses. The Gaussian Process Subset Scan enables early detection of emerging patterns in spatio-te…

Anomaly DetectionBIG-bench Machine Learning

Few-Shot Learning with Uncertainty-based Quadruplet Selection for Interference Classification in GNSS Data

2024-02-09 · Felix Ott, Lucas Heublein, Nisha Lakshmana Raichur, Tobias Feigl 외

Jamming devices pose a significant threat by disrupting signals from the global navigation satellite system (GNSS), compromising the robustness of accurate positioning. Detecting anomalies in frequency snapshots is cruci…

Few-Shot Learning