paper-with-me

홈 › Papers

Learning to Think in Physics: Breaking Shortcut Learning in Scientific Diffusion via Representation Alignment

2026-05-20 · Haozhe Jia, Pengyu Yin, Wenshuo Chen, Shaofeng Liang, Lei Wang, Bowen Tian, Xiucheng Wang, Nanqian Jia, Yutao Yue arxiv

Physics-informed diffusion models typically enforce PDE constraints only on final outputs, leaving intermediate representations unconstrained and prone to shortcut learning under shifted boundary conditions. We introduce REPA-P, a teacher-free, architecture-agnostic framework that aligns intermediate features with physical states using first-principles residuals. REPA-P attaches lightweight $1{\times}1$ projection heads to selected layers, decodes hidden activations into physical quantities, and applies PDE residual losses during training. These heads are discarded at inference, introducing zero overhead. Across four PDE tasks, including Darcy flow, topology optimization, electrostatic potential, and turbulent channel flow, REPA-P accelerates convergence by up to $2{\times}$, reduces physics residuals by up to $66.4\%$, and improves out-of-distribution robustness by up to $49.3\%$, with consistent gains on both U-Net and Diffusion Transformer backbones. Ablations show that supervising a small set of intermediate layers captures most benefits and complements output-level physics losses. Code is available at https://github.com/Hxxxz0/REPA-P.

📄 PDF Abstract BibTeX arXiv:2605.20780

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

S1-VL: Scientific Multimodal Reasoning Model with Thinking-with-Images

2026-04-23 · Qingxiao Li, Lifeng Xu, QingLi Wang, Yudong Bai 외 arxiv

We present S1-VL, a multimodal reasoning model for scientific domains that natively supports two complementary reasoning paradigms: Scientific Reasoning, which relies on structured chain-of-thought, and Thinking-with-Ima…

Reinforcement LearningMultimodal Reasoning

Statistical physics of unsupervised learning with prior knowledge in neural networks

2019-11-06 · Tianqi Hou, Haiping Huang

Integrating sensory inputs with prior beliefs from past experiences in unsupervised learning is a common and fundamental characteristic of brain or artificial neural computation. However, a quantitative role of prior kno…

The Perception-Physics Paradox: Probing Scientific Alignment with TC-Bench

2026-05-23 · Dingling Yao, Andrea Polesello, Adeel Pervez, Caroline Muller 외 arxiv

While Vision Foundation Models (VFMs) excel at predictive tasks on satellite imagery, their performance can arise from visual correlations rather than underlying structural invariants, making even perception-based out-of…

Representation Learning

Rethinking SO(3)-equivariance with Bilinear Tensor Networks

2023-03-20 · Chase Shimmin, Zhelun Li, Ema Smith

Many datasets in scientific and engineering applications are comprised of objects which have specific geometric structure. A common example is data which inhabits a representation of the group SO$(3)$ of 3D rotations: sc…

Tensor Networks

Symmetry-Breaking De Novo Crystal Generation via Markovian Jump Diffusion

2026-08-13 · Van Khoa Nguyen, Alexandros Kalousis arxiv

Generating crystals has recently attracted significant interest due to their broad applications in materials science. However, existing generative models struggle to produce complete crystallographic specifications, limi…