paper-with-me

Papers

Dynamic sampling bias and overdispersion induced by skewed offspring distributions

2021-03-10 · Takashi Okada, Oskar Hallatschek

Natural populations often show enhanced genetic drift consistent with a strong skew in their offspring number distribution. The skew arises because the variability of family sizes is either inherently strong or amplified by population expansions, leading to so-called `jackpot' events. The resulting allele frequency fluctuations are large and, therefore, challenge standard models of population genetics, which assume sufficiently narrow offspring distributions. While the neutral dynamics backward in time can be readily analyzed using coalescent approaches, we still know little about the effect of broad offspring distributions on the dynamics forward in time, especially with selection. Here, we employ an exact asymptotic analysis combined with a scaling hypothesis to demonstrate that over-dispersed frequency trajectories emerge from the competition of conventional forces, such as selection or mutations, with an emerging time-dependent sampling bias against the minor allele. The sampling bias arises from the characteristic time-dependence of the largest sampled family size within each allelic type. Using this insight, we establish simple scaling relations for allele frequency fluctuations, fixation probabilities, extinction times, and the site frequency spectra that arise when offspring numbers are distributed according to a power law $~n^{-(1+\alpha)}$. To demonstrate that this coarse-grained model captures a wide variety of non-equilibrium dynamics, we validate our results in traveling waves, where the phenomenon of 'gene surfing' can produce any exponent $1<\alpha <2$. We argue that the concept of a dynamic sampling bias is useful generally to develop both intuition and statistical tests for the unusual dynamics of populations with skewed offspring distributions, which can confound commonly used tests for selection or demographic history.

📄 PDF Abstract BibTeX arXiv:2103.05852

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Traversing the noise of dynamic mini-batch sub-sampled loss functions: A visual guide

2019-03-20 · Dominic Kafka, Daniel Wilke

Mini-batch sub-sampling in neural network training is unavoidable, due to growing data demands, memory-limited computational resources such as graphical processing units (GPUs), and the dynamics of on-line learning. In t…

Stochastic Optimization

Effects of sampling skewness of the importance-weighted risk estimator on model selection

2018-04-19 · Wouter M. Kouw, Marco Loog

Importance-weighting is a popular and well-researched technique for dealing with sample selection bias and covariate shift. It has desirable characteristics such as unbiasedness, consistency and low computational complex…

Model SelectionSelection bias

Federated Unlearning Model Recovery in Data with Skewed Label Distributions

2024-12-18 · Xinrui Yu, Wenbin Pei, Bing Xue, Qiang Zhang

In federated learning, federated unlearning is a technique that provides clients with a rollback mechanism that allows them to withdraw their data contribution without training from scratch. However, existing research ha…

DenoisingFederated Learning

Fast sampling from constrained spaces using the Metropolis-adjusted Mirror Langevin algorithm

2023-12-14 · Vishwak Srinivasan, Andre Wibisono, Ashia Wilson

We propose a new method called the Metropolis-adjusted Mirror Langevin algorithm for approximate sampling from distributions whose support is a compact and convex set. This algorithm adds an accept-reject filter to the M…

Self-Organizing Maps of Unbiased Ligand-Target Binding Pathways and Kinetics

2024-09-19 · Lara Callea, Camilla Caprai, Laura Bonati, Toni Giorgino 외

The interpretation of ligand-target interactions at atomistic resolution is central to most efforts in computational drug discovery and optimization. However, the highly dynamic nature of protein targets, as well as poss…

Drug DesignDrug DiscoverySpecificity