paper-with-me

홈 › Papers

Non-Identifiable Pedigrees and a Bayesian Solution

2016-02-26

Some methods aim to correct or test for relationships or to reconstruct the pedigree, or family tree. We show that these methods cannot resolve ties for correct relationships due to identifiability of the pedigree likelihood which is the probability of inheriting the data under the pedigree model. This means that no likelihood-based method can produce a correct pedigree inference with high probability. This lack of reliability is critical both for health and forensics applications. In this paper we present the first discussion of multiple typed individuals in non-isomorphic pedigrees, $\mathcal{P}$ and $\mathcal{Q}$, where the likelihoods are non-identifiable, $Pr[G~|~\mathcal{P},\theta] = Pr[G~|~\mathcal{Q},\theta]$, for all input data $G$ and all recombination rate parameters $\theta$. While there were previously known non-identifiable pairs, we give an example having data for multiple individuals. Additionally, deeper understanding of the general discrete structures driving these non-identifiability examples has been provided, as well as results to guide algorithms that wish to examine only identifiable pedigrees. This paper introduces a general criteria for establishing whether a pair of pedigrees is non-identifiable and two easy-to-compute criteria guaranteeing identifiability. Finally, we suggest a method for dealing with non-identifiable likelihoods: use Bayes rule to obtain the posterior from the likelihood and prior. We propose a prior guaranteeing that the posterior distinguishes all pairs of pedigrees. Shortened version published as: B. Kirkpatrick. Non-identifiable pedigrees and a Bayesian solution. Int. Symp. on Bioinformatics Res. and Appl. (ISBRA), 7292:139-152 2012.

📄 PDF Abstract BibTeX arXiv:1602.08183

Code (0)

등록된 구현이 없습니다.

Similar Papers 제목 키워드 기반

Bayesian Pedigree Analysis using Measure Factorization

2012-12-01 · NeurIPS 2012 12 · Bonnie Kirkpatrick, Alexandre Bouchard-Côté

Pedigrees, or family trees, are directed graphs used to identify sites of the genome that are correlated with the presence or absence of a disease. With the advent of genotyping and sequencing technologies, there has be…

Efficient Reconstruction of Stochastic Pedigrees

2020-05-08 · Younhun Kim, Elchanan Mossel, Govind Ramnarayan, Paxton Turner

We introduce a new algorithm called {\sc Rec-Gen} for reconstructing the genealogy or \textit{pedigree} of an extant population purely from its genetic data. We justify our approach by giving a mathematical proof of the …

Blang: Bayesian declarative modelling of general data structures and inference via algorithms based on distribution continua

2019-12-22 · Alexandre Bouchard-Côté, Kevin Chern, Davor Cubranic, Sahand Hosseini 외

Consider a Bayesian inference problem where a variable of interest does not take values in a Euclidean space. These "non-standard" data structures are in reality fairly common. They are frequently used in problems involv…

Bayesian Inference

Learning Bayesian and Markov Networks with an Unreliable Oracle

2026-03-10 · Juha Harviainen, Pekka Parviainen, Vidya Sagar Sharma arxiv

We study constraint-based structure learning of Markov networks and Bayesian networks in the presence of an unreliable conditional independence oracle that makes at most a bounded number of errors. For Markov networks, w…

Recent Advances in Algebraic Geometry and Bayesian Statistics

2022-11-18 · Sumio Watanabe

This article is a review of theoretical advances in the research field of algebraic geometry and Bayesian statistics in the last two decades. Many statistical models and learning machines which contain hierarchical struc…