paper-with-me

홈 › Papers

NOVA: Fundamental Limits of Knowledge Discovery Through AI

2026-05-12 · Salman Avestimehr, Ken Duffy, Muriel Médard arxiv

Can AI systems discover genuinely new knowledge through iterative self improvement, and if so, at what cost? We introduce the NOVA framework, which models the common ``generate, verify, accumulate, retrain'' loop as an adaptive sampling process over a knowledge space. We identify sufficient conditions under which accumulated genuine knowledge eventually covers a finite domain, and show how their violations produce distinct failure modes: contamination, forgetting, exploration failure, and acceptance failure. We then analyze imperfect verification and identify a contamination trap: as easy-to-find knowledge is exhausted, the model mass assigned to new valid artifacts shrinks, so even small false-positive rates can cause invalid artifacts to enter the knowledge base faster than genuine discoveries. We clarify that Good--Turing estimation is a local batch-diversity diagnostic, not an estimator of the historically undiscovered valid mass that governs long-term discovery. Under a separate tail-equivalence assumption relating the model's effective discovery distribution to a Zipf law with exponent $α>1$, we prove that the cumulative generation cost required to obtain $D$ distinct genuine discoveries satisfies $R_{\mathrm{cum}}(D)=Θ(c_{\mathrm{gen}}D^α)$, where $c_{\mathrm{gen}}$ is the per-candidate generation cost. This scaling law quantifies asymptotic diminishing returns as the discovery frontier advances. Finally, we formalize human amplification through guidance, generation, and verification, explaining why expert input is most valuable near autonomous exploration barriers.

📄 PDF Abstract BibTeX arXiv:2605.15219

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Fundamental Limits of Deep Learning-Based Binary Classifiers Trained with Hinge Loss

2023-09-13 · Tilahun M. Getu, Georges Kaddoum

Although deep learning (DL) has led to several breakthroughs in many disciplines as diverse as chemistry, computer science, electrical engineering, mathematics, medicine, neuroscience, and physics, a comprehensive unders…

Electrical Engineering

Fundamental Limits of Prediction, Generalization, and Recursion: An Entropic-Innovations Perspective

2020-01-12 · Song Fang, Quanyan Zhu

In this paper, we examine the fundamental performance limits of prediction, with or without side information. More specifically, we derive generic lower bounds on the $\mathcal{L}_p$ norms of the prediction errors that a…

Predictionvalid

AlphaGo Moment for Model Architecture Discovery

2025-07-24 · Yixiu Liu, Yang Nan, Weixian Xu, Xiangkun Hu 외 arxiv

While AI systems demonstrate exponentially improving capabilities, the pace of AI research itself remains linearly bounded by human cognitive capacity, creating an increasingly severe development bottleneck. We present A…

Neural Architecture Search

Hierarchical Physics-Embedded Learning for Prediction and Discovery in Spatiotemporal Dynamical Systems

2025-10-29 · Xizhe Wang, Xiaobin Song, Qingshan Jia, Hao Sun 외 arxiv

Modeling complex spatiotemporal dynamics, particularly in far-from-equilibrium systems, remains a grand challenge in science. The governing partial differential equations (PDEs) for these systems are often intractable to…

The Clinical Trials Puzzle: How Network Effects Limit Drug Discovery

2023-01-25 · Kishore Vasan, Deisy Gysi, Albert-Laszlo Barabasi

The depth of knowledge offered by post-genomic medicine has carried the promise of new drugs, and cures for multiple diseases. To explore the degree to which this capability has materialized, we extract meta-data from 35…

Drug Discovery