paper-with-me

Papers

Initial Model Incorporation for Deep Learning FWI: Pretraining or Denormalization?

2025-06-05 · Ruihua Chen, Bangyu Wu, Meng Li, Kai Yang

Subsurface property neural network reparameterized full waveform inversion (FWI) has emerged as an effective unsupervised learning framework, which can invert stably with an inaccurate starting model. It updates the trainable neural network parameters instead of fine-tuning on the subsurface model directly. There are primarily two ways to embed the prior knowledge of the initial model into neural networks, that is, pretraining and denormalization. Pretraining first regulates the neural networks' parameters by fitting the initial velocity model; Denormalization directly adds the outputs of the network into the initial models without pretraining. In this letter, we systematically investigate the influence of the two ways of initial model incorporation for the neural network reparameterized FWI. We demonstrate that pretraining requires inverting the model perturbation based on a constant velocity value (mean) with a two-stage implementation. It leads to a complex workflow and inconsistency of objective functions in the two-stage process, causing the network parameters to become inactive and lose plasticity. Experimental results demonstrate that denormalization can simplify workflows, accelerate convergence, and enhance inversion accuracy compared with pretraining.

📄 PDF Abstract BibTeX arXiv:2506.05484

Code (0)

등록된 구현이 없습니다.

Tasks

Deep Learning

Similar Papers 제목 키워드 기반

Pretraining the Noisy Channel Model for Task-Oriented Dialogue

2021-03-18 · Qi Liu, Lei Yu, Laura Rimell, Phil Blunsom

Direct decoding for task-oriented dialogue is known to suffer from the explaining-away effect, manifested in models that prefer short and generic responses. Here we argue for the use of Bayes' theorem to factorize the di…

End-To-End Dialogue Modelling

Never Train from Scratch: Fair Comparison of Long-Sequence Models Requires Data-Driven Priors

2023-10-04 · Ido Amos, Jonathan Berant, Ankit Gupta

Modeling long-range dependencies across sequences is a longstanding goal in machine learning and has led to architectures, such as state space models, that dramatically outperform Transformers on long sequences. However,…

DenoisingState Space Models

On Training Derivative-Constrained Neural Networks

2023-10-02 · KaiChieh Lo, Daniel Huang

We refer to the setting where the (partial) derivatives of a neural network's (NN's) predictions with respect to its inputs are used as additional training signal as a derivative-constrained (DC) NN. This situation is co…

Face Deblurring Based on Separable Normalization and Adaptive Denormalization

2021-12-18 · Xian Zhang, Hao Zhang, Jiancheng Lv, Xiaojie Li

Face deblurring aims to restore a clear face image from a blurred input image with more explicit structure and facial details. However, most conventional image and face deblurring methods focus on the whole generated ima…

DeblurringFace ParsingSSIM

LocCa: Visual Pretraining with Location-aware Captioners

2024-03-28 · Bo Wan, Michael Tschannen, Yongqin Xian, Filip Pavetic 외

Image captioning has been shown as an effective pretraining method similar to contrastive pretraining. However, the incorporation of location-aware information into visual pretraining remains an area with limited researc…

DecoderImage Captioning