Inference for Large Panel Data with Many Covariates
This paper proposes a novel testing procedure for selecting a sparse set of covariates that explains a large dimensional panel. Our selection method provides correct false detection control while having higher power than existing approaches. We develop the inferential theory for large panels with many covariates by combining post-selection inference with a novel multiple testing adjustment. Our data-driven hypotheses are conditional on the sparse covariate selection. We control for family-wise error rates for covariate discovery for large cross-sections. As an easy-to-use and practically relevant procedure, we propose Panel-PoSI, which combines the data-driven adjustment for panel multiple testing with valid post-selection p-values of a generalized LASSO, that allows us to incorporate priors. In an empirical study, we select a small number of asset pricing factors that explain a large cross-section of investment strategies. Our method dominates the benchmarks out-of-sample due to its better size and power.
Code (0)
등록된 구현이 없습니다.
Tasks
validSimilar Papers 제목 키워드 기반
Double Machine Learning for Static Panel Data with Instrumental Variables: New Method and Applications
Panel data methods are widely used in empirical analysis to address unobserved heterogeneity, but causal inference remains challenging when treatments are endogenous and confounding variables high-dimensional and potenti…
Causal InferenceBounds on Average Effects in Discrete Choice Panel Data Models
In discrete choice panel data, the estimation of average effects is crucial for quantifying the effect of covariates, and for policy evaluation and counterfactual analysis. This task is challenging in short panels with i…
counterfactualvalidRoot-n-consistent Conditional ML estimation of dynamic panel logit models with fixed effects
In this paper we first propose a root-n-consistent Conditional Maximum Likelihood (CML) estimator for all the common parameters in the panel logit AR(p) model with strictly exogenous covariates and fixed effects. Our CML…
Data-driven model selection within the matrix completion method for causal panel data models
Matrix completion estimators are employed in causal panel data models to regulate the rank of the underlying factor model using nuclear norm minimization. This convex optimization problem enables concurrent regularizatio…
Matrix CompletionModel Selectionparameter estimationvalid+1Nickell Meets Stambaugh: A Tale of Two Biases in Panel Predictive Regressions
In panel predictive regressions with persistent covariates, coexistence of the Nickell bias and the Stambaugh bias imposes challenges for hypothesis testing. This paper introduces a new estimator, the IVX-X-Jackknife (IV…