High Dimensional Binary Choice Model with Unknown Heteroskedasticity or Instrumental Variables
This paper proposes a new method for estimating high-dimensional binary choice models. The model we consider is semiparametric, placing no distributional assumptions on the error term, allowing for heteroskedastic errors, and permitting endogenous regressors. Our proposed approaches extend the special regressor estimator originally proposed by Lewbel (2000). This estimator becomes impractical in high-dimensional settings due to the curse of dimensionality associated with high-dimensional conditional density estimation. To overcome this challenge, we introduce an innovative data-driven dimension reduction method for nonparametric kernel estimators, which constitutes the main innovation of this work. The method combines distance covariance-based screening with cross-validation (CV) procedures, rendering the special regressor estimation feasible in high dimensions. Using the new feasible conditional density estimator, we address the variable and moment (instrumental variable) selection problems for these models. We apply penalized least squares (LS) and Generalized Method of Moments (GMM) estimators with a smoothly clipped absolute deviation (SCAD) penalty. A comprehensive analysis of the oracle and asymptotic properties of these estimators is provided. Monte Carlo simulations are employed to demonstrate the effectiveness of our proposed procedures in finite sample scenarios.
Code (0)
등록된 구현이 없습니다.
Tasks
Density EstimationDimensionality ReductionVariable SelectionSimilar Papers 제목 키워드 기반
A Bayesian Perspective on the Maximum Score Problem
This paper presents a Bayesian inference framework for a linear index threshold-crossing binary choice model that satisfies a median independence restriction. The key idea is that the model is observationally equivalent …
Bayesian InferenceRobust Inference of Manifold Density and Geometry by Doubly Stochastic Scaling
The Gaussian kernel and its traditional normalizations (e.g., row-stochastic) are popular approaches for assessing similarities between data points. Yet, they can be inaccurate under high-dimensional noise, especially if…
Dynamic Pricing in High-dimensions
We study the pricing problem faced by a firm that sells a large number of products, described via a wide range of features, to customers that arrive over time. Customers independently make purchasing decisions according …
Vocal Bursts Intensity PredictionStandard Errors for Panel Data Models with Unknown Clusters
This paper develops a new standard-error estimator for linear panel data models. The proposed estimator is robust to heteroskedasticity, serial correlation, and cross-sectional correlation of unknown forms. The serial co…
Diffusion Models Adapt to Low-Dimensional Structure Under Flexible Coefficient Choices
Diffusion models are known to exploit unknown low-dimensional structure to accelerate sampling. However, existing convergence theory under low-dimensional data structure has largely focused on update rules with narrowly …