paper-with-me

홈 › Papers

DN-Hypo-Pipeline: An AI-Driven Workflow for Generating Hypotheses using Large Language Models and Scientific Explanations

2026-06-07 · Lei Lin, Ronghao Wang, Chunbao Zhou, Jue Wang, Yangang Wang arxiv

Modern artificial intelligence excels at prediction but cannot explain. From large language models to AI-for-science systems, today's machines answer what by recombining patterns already present in the human literature, yet they cannot reason out why a phenomenon must arise from underlying principles even though explanation, not prediction, lies at the heart of scientific discovery. Here we ask whether the structure of scientific explanation can be operationalized to guide how a machine generates hypotheses. We introduce DN-Hypo-Pipeline, a hypothesis-generation framework that adopts a layered, explanation-theoretic scaffold: Hempel's deductive-nomological (DN) model supplies the output form and deductive validity of a hypothesis, Salmon's causal-process account supplies an organizing constraint on where to search for the governing laws, and Armstrong's view of laws as relations between universals supplies the bridge from a phenomenon's constituent processes to the laws that may be associated with it. Rather than searching the space of what has been written, the framework searches the space of what principles govern a phenomenon: given an explanandum, it abstracts the universals instantiated in the phenomenon's formation process, retrieves the laws relating those universals, and deductively reconstructs a new, testable explanation. Evaluated in data-science modeling and judged by both LLMs and human experts, hypotheses generated through this principled reasoning significantly outperform those from direct prompting. Crucially, we translated the two highest-scoring hypotheses into novel algorithms one that reduces the Transformer's theoretical complexity with only minimal performance loss, and another that achieves competitive accuracy with substantially fewer parameters.

📄 PDF Abstract BibTeX arXiv:2606.08532

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Agentic Discovery: Closing the Loop with Cooperative Agents

2025-10-15 · J. Gregory Pauloski, Kyle Chard, Ian T. Foster arxiv

As data-driven methods, artificial intelligence (AI), and automated workflows accelerate scientific tasks, we see the rate of discovery increasingly limited by human decision-making tasks such as setting objectives, gene…

Continuous Knowledge Metabolism: Generating Scientific Hypotheses from Evolving Literature

2026-04-14 · Jinkai Tao, Yubo Wang, Xiaoyu Liu, Menglin Yang arxiv

Identifying promising research directions in fast-moving subareas is one of the most cognitively expensive tasks in modern AI research. Existing LLM-driven scientific discovery systems are typically limited to one-shot p…

Automatically Generating Hard Math Problems from Hypothesis-Driven Error Analysis

2026-04-06 · Jiayu Fu, Mourad Heddaya, Chenhao Tan arxiv

Numerous math benchmarks exist to evaluate LLMs' mathematical capabilities. However, most involve extensive manual effort and are difficult to scale. Consequently, they cannot keep pace with LLM development or easily pro…

The Qualitative Laboratory: Theory Prototyping and Hypothesis Generation with Large Language Models

2025-11-25 · Hugues Draelants arxiv

A central challenge in social science is to generate rich qualitative hypotheses about how diverse social groups might interpret new information. This article introduces and illustrates a novel methodological approach fo…

HARPA: A Testability-Driven, Literature-Grounded Framework for Research Ideation

2025-10-01 · Rosni Vasu, Peter Jansen, Pao Siangliulue, Cristina Sarasua 외 arxiv

While there has been a surge of interest in automated scientific discovery (ASD), especially with the emergence of LLMs, it remains challenging for tools to generate hypotheses that are both testable and grounded in the …