paper-with-me

Papers

Can LLMs extract scientific consensus? A case study in high-temperature superconductivity

2026-05-26 · Mouyang Cheng, Wenhao He, Zhuotao Jin, Bowen Yu, Ju Li, Boris Kozinsky, Yao Wang, Pavel Volkov, Liangzi Deng, Ching-Wu Chu, Xiao-Gang Wen, Mingda Li arxiv

Scientific knowledge is increasingly dispersed across vast and heterogeneous scientific literature, where important claims are often implicit, evolving, and internally debated. While large language models (LLMs) have shown impressive performance in information extraction and summarization, their ability to recover latent scientific consensus remains unclear. Here, we investigate this problem in the context of high-temperature superconductivity (HTS), a long-standing and highly debated topic in condensed matter physics, as a challenging testbed. Using near 18,000 highly-cited publications over the past seven decades, we construct a structured knowledge graph linking competing superconducting mechanisms, material families, evidential modalities, and citation relations. We find that LLM-extracted representations recover coherent and physically interpretable structures, including family-dependent mechanism profiles, evidence-specific correlations, and citation-mediated temporal evolution of scientific beliefs. Ablation studies on LLM further show that the global structure remains robust across prompting, decoding, and model variations. Our results suggest that LLMs can indeed serve as scalable tools for deciphering scientific knowledge in domains characterized by competing interpretations and evolving knowledge.

📄 PDF Abstract BibTeX arXiv:2606.07570

Code (0)

등록된 구현이 없습니다.

Tasks

Information Extraction

Similar Papers 제목 키워드 기반

Self-prompting and cross-model consensus enable reproducible data extraction from scientific literature with large language models

2026-08-19 · Valentin Romanov, Monique Bax, Steven Niederer arxiv

Accurately extracting nuanced, contextualized data from research articles is laborious and time intensive. Here, we investigate the performance of frontier, browser-based large language models (LLMs) to extract highly co…

Large language models eroding science understanding: an experimental study

2026-04-28 · Harry Collins, Hartmut Grote, Paul Newbury, Patrick Sutton 외 arxiv

This paper is under review in AI and Ethics This study examines whether large language models (LLMs) can reliably answer scientific questions and demonstrates how easily they can be influenced by fringe scientific materi…

TurQUaz at CheckThat! 2025: Debating Large Language Models for Scientific Web Discourse Detection

2025-07-26 · Tarık Saraç, Selin Mergen, Mucahid Kutlu arxiv

In this paper, we present our work developed for the scientific web discourse detection task (Task 4a) of CheckThat! 2025. We propose a novel council debate method that simulates structured academic discussions among mul…

Scientific discovery as meta-optimization: a combinatorial optimization case study

2026-06-25 · Yuan-Hang Zhang, Chesson Sipling, Massimiliano Di Ventra arxiv

Scientific discovery is fundamentally an optimization problem, defined by a vast "state space" of theories and experiments, and an evaluation criterion based on quality, novelty, and validity. Large language models (LLMs…

Digital Pathway Curation (DPC): a pipeline able to assess the reproducibility, consensus and accuracy in biomedical search retrieval by comparing Gemini, PubMed, and Scientific Reviewers

2025-05-02 · Flavio Lichtenstein, Daniel Alexandre de Souza, Carlos Eduardo Madureira Trufen, Victor Wendel da Silva Gonçalves 외

A scientific study begins with a central question, and search engines like PubMed are the first tools for retrieving knowledge and understanding the current state of the art. Large Language Models (LLMs) have been used i…