Independence in Integrated Population Models
Integrated population models (IPMs) combine multiple ecological data types such as capture-mark-recapture histories, reproduction surveys, and population counts into a single statistical framework. In such models, each data type is generated by a probabilistic submodel, and an assumption of independence between the different data types is usually made. The fact that the same biological individuals can contribute to multiple data types has been perceived as affecting their independence, and several studies have even investigated IPM robustness in this scenario. However, what matters from a statistical perspective is probabilistic independence: the joint probability of observing all data is equal to the product of the likelihoods of the various datasets. Contrary to a widespread perception, probabilistic non-independence does not automatically result from collecting data on the same physical individuals. Conversely, while there can be good reasons for non-independence of IPM submodels arising from sharing of individuals between data types, these relations do not seem to be included in IPMs whose robustness is being investigated. Furthermore, conditional rather than true independence is sometimes assumed. In this conceptual paper, I survey the various independence concepts used in IPMs, try to make sense of them by getting back to first principles in toy models, and show that it is possible to obtain probabilistic independence (or near-independence) despite two or three data types collected on the same set of biological individuals. I then revisit recommendations pertaining to component data collection and IPM robustness checks, and provide some suggestions to bridge the current gap between individual-level IPMs and their population-level approximations using composite likelihoods.
Code (0)
등록된 구현이 없습니다.
Methods 이 논문이 사용한 방법론
Similar Papers 제목 키워드 기반
USP: an independence test that improves on Pearson's chi-squared and the $G$-test
We present the $U$-Statistic Permutation (USP) test of independence in the context of discrete data displayed in a contingency table. Either Pearson's chi-squared test of independence, or the $G$-test, are typically used…
Conditional Independence Test Based on Transport Maps
Testing conditional independence between two random vectors given a third is a fundamental and challenging problem in statistics, particularly in multivariate nonparametric settings due to the complexity of conditional s…
Maximal Social Welfare Relations on Infinite Populations Satisfying Permutation Invariance
We study social welfare relations (SWRs) on an infinite population. Our main result is a new characterization of a utilitarian SWR as the \emph{largest} SWR (in terms of subset when the weak relation is viewed as a set o…
RelationA Framework for Inferring Causality from Multi-Relational Observational Data using Conditional Independence
The study of causality or causal inference - how much a given treatment causally affects a given outcome in a population - goes way beyond correlation or association analysis of variables, and is critical in making sound…
Causal InferenceStable Heterogeneous Treatment Effect Estimation across Out-of-Distribution Populations
Heterogeneous treatment effect (HTE) estimation is vital for understanding the change of treatment effect across individuals or subgroups. Most existing HTE estimation methods focus on addressing selection bias induced b…
counterfactualHeterogeneous Treatment Effect EstimationRepresentation LearningSelection bias