paper-with-me

홈 › Papers

CDRL: Certification-Driven Reinforcement Learning for Neutrino Flavor Model Discovery

2026-08-21 · Piyush Jha, Jake Rudolph, Victoria Knapp-Pérez, Max Fieg, Aishik Ghosh, Vijay Ganesh arxiv

Many scientific discovery problems require searching combinatorial hypothesis spaces under complex domain constraints. Reinforcement learning (RL) offers a promising approach, but existing methods rely on scalar rewards that provide limited information about why candidate solutions fail, leading agents to repeatedly explore invalid regions. We introduce Certification-Driven Reinforcement Learning (CDRL), a framework that leverages structured feedback from symbolic reasoning tools. When a candidate violates domain constraints, these tools produce certificates identifying the actions responsible for failure. CDRL converts these certificates into reusable constraints that eliminate classes of invalid solutions and guide exploration toward valid regions. We evaluate CDRL on neutrino flavor model discovery in theoretical particle physics, where the hypothesis space exceeds $10^{26}$ possible models, and compare it with the state-of-the-art RL approach previously used for this task. Across three theory spaces, CDRL achieves up to 1.95$\times$ higher valid model rates and up to 6.33$\times$ higher neutrino model rates while evaluating up to 4$\times$ fewer candidates. We further extract 40 interpretable rules from search trajectories using a post-hoc decision-tree framework and show that reusing them as soft constraints yields gains of up to 2$\times$ in valid model rates and 3$\times$ in neutrino model discovery across all three theory spaces. These results suggest that CDRL uncovers reusable structure in combinatorial search spaces and provides a general framework for scientific model discovery.

📄 PDF Abstract BibTeX arXiv:2608.20686

Code (0)

등록된 구현이 없습니다.

Tasks

Reinforcement Learning

Similar Papers 제목 키워드 기반

Application of Neural Networks for the Reconstruction of Supernova Neutrino Energy Spectra Following Fast Neutrino Flavor Conversions

2024-01-30 · Sajad Abbar, Meng-Ru Wu, Zewei Xiong

Neutrinos can undergo fast flavor conversions (FFCs) within extremely dense astrophysical environments such as core-collapse supernovae (CCSNe) and neutron star mergers (NSMs). In this study, we explore FFCs in a \emph{m…

2D Convolutional Neural Network for Event Reconstruction in IceCube DeepCore

2023-07-31 · J. H. Peterson, M. Prado Rodriguez, K. Hanson

IceCube DeepCore is an extension of the IceCube Neutrino Observatory designed to measure GeV scale atmospheric neutrino interactions for the purpose of neutrino oscillation studies. Distinguishing muon neutrinos from oth…

Exploring the flavor structure of quarks and leptons with reinforcement learning

2023-04-27 · Satsuki Nishimura, Coh Miyao, Hajime Otsuka

We propose a method to explore the flavor structure of quarks and leptons with reinforcement learning. As a concrete model, we utilize a basic value-based algorithm for models with $U(1)$ flavor symmetry. By training neu…

reinforcement-learningReinforcement Learning

Exploring the flavor structure of leptons via diffusion models

2025-03-27 · Satsuki Nishimura, Hajime Otsuka, Haruki Uchiyama

We propose a method to explore the flavor structure of leptons using diffusion models, which are known as one of generative artificial intelligence (generative AI). We consider a simple extension of the Standard Model wi…

Transfer Learning

Solar neutrino limit on axions and keV-mass bosons

2008-07-18 · Paolo Gondolo, Georg Raffelt

The all-flavor solar neutrino flux measured by the Sudbury Neutrino Observatory (SNO) constrains nonstandard energy losses to less than about 10% of the Sun's photon luminosity, superseding a helioseismological argument …