paper-with-me

홈 › Papers

Unlearning as Ablation: Toward a Falsifiable Benchmark for Generative Scientific Discovery

2025-08-25 · Robert Yang arxiv

Bold claims about AI's role in science-from "AGI will cure all diseases" to promises of radically accelerated discovery-raise a central epistemic question: do large language models (LLMs) truly generate new knowledge, or do they merely remix memorized fragments? We propose unlearning-as-ablation as a falsifiable probe of constructive scientific discovery. The idea is to systematically remove a target result together with its forget-closure (supporting lemmas, paraphrases, and multi-hop entailments) and then evaluate whether the model can re-derive the result from only permitted axioms and tools. Success would indicate generative capability beyond recall; failure would expose current limits. Unlike prevailing motivations for unlearning-privacy, copyright, or safety-our framing repositions it as an epistemic probe for AI-for-Science. We outline a minimal pilot in mathematics and algorithms to illustrate feasibility, and sketch how the same approach could later be extended to domains such as physics or chemistry. This is a position paper: our contribution is conceptual and methodological, not empirical. We aim to stimulate discussion on how principled ablation tests could help distinguish models that reconstruct knowledge from those that merely retrieve it, and how such probes might guide the next generation of AI-for-Science benchmarks.

📄 PDF Abstract BibTeX arXiv:2508.17681

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Can Scientific Claims Be Removed from Large Language Models? A Systematic Evaluation of Claim-Level Unlearning

2026-08-21 · Snigdha Paul, Manasi Patwardhan, Arman Cohan arxiv

Language models (LMs) are trained on static scientific corpora, whereas scientific knowledge continuously evolves through correction and revision. Scientific claims encoded within these models may later become retracted,…

Can Large Language Models Reinvent Foundational Algorithms?

2026-04-07 · Jian Zhao, Haoren Luo, Yu Wang, Yuhan Cao 외 arxiv

LLMs have shown strong potential to advance scientific discovery. Whether they possess the capacity for foundational innovation, however, remains an open question. In this work, we focus on a prerequisite for foundationa…

Reinforcement Learning

Protecting Privacy in Multimodal Large Language Models with MLLMU-Bench

2024-10-29 · Zheyuan Liu, Guangyao Dou, Mengzhao Jia, Zhaoxuan Tan 외

Generative models such as Large Language Models (LLM) and Multimodal Large Language models (MLLMs) trained on massive web corpora can memorize and disclose individuals' confidential and private data, raising legal and et…

Language ModelingLanguage ModellingLarge Language ModelMachine Unlearning+1

Controllable Unlearning for Image-to-Image Generative Models via $\varepsilon$-Constrained Optimization

2024-08-03 · Xiaohua Feng, Chaochao Chen, Yuyuan Li, Li Zhang

While generative models have made significant advancements in recent years, they also raise concerns such as privacy breaches and biases. Machine unlearning has emerged as a viable solution, aiming to remove specific tra…

Machine Unlearningvalid

On the Limitations and Prospects of Machine Unlearning for Generative AI

2024-08-01 · Shiji Zhou, Lianzhe Wang, Jiangnan Ye, Yongliang Wu 외

Generative AI (GenAI), which aims to synthesize realistic and diverse data samples from latent variables or other data modalities, has achieved remarkable results in various domains, such as natural language, images, aud…

EthicsMachine Unlearning