To make high-quality research more accessible and easier to explore.

Fields:
11 results

Inference for Heterogeneous Effects using Low-Rank Estimation of Factor Slopes

The Review of Economics and Statistics 2026
We study a panel data model with heterogeneous effects, allowing slopes to vary across individuals and time. To reduce dimensionality, we assume these slopes follow a factor structure, so slope matrices can be estimated via low-rank regularized regression. We propose a multi-step estimation procedure incorporating sample splitting and partialing-out to enable valid inference after penalized estimation. We establish the asymptotic normality of the resulting estimator, facilitating inference for individualtime- specific effects and their cross-sectional averages. The method’s performance is illustrated through simulations and an empirical application.

Inference with Dependent Data in Accounting and Finance Applications

Journal of Accounting Research 2018 56(4), 1139-1203 open access
We review developments in conducting inference for model parameters in the presence of intertemporal and cross‐sectional dependence with an emphasis on panel data applications. We review the use of heteroskedasticity and autocorrelation consistent (HAC) standard error estimators, which include the standard clustered and multiway clustered estimators, and discuss alternative sample‐splitting inference procedures, such as the Fama–Macbeth procedure, within this context. We outline pros and cons of the different procedures. We then illustrate the properties of the discussed procedures within a simulation experiment designed to mimic the type of firm‐level panel data that might be encountered in accounting and finance applications. Our conclusion, based on theoretical properties and simulation performance, is that sample‐splitting procedures with suitably chosen splits are the most likely to deliver robust inferential statements with approximately correct coverage properties in the types of large, heterogeneous panels many researchers are likely to face.

An IV Model of Quantile Treatment Effects

Econometrica 2005 73(1), 245-261 open access
The ability of quantile regression models to characterize the heterogeneous impact of variables on different points of an outcome distribution makes them appealing in many economic applications. However, in observational studies, the variables of interest (e.g., education, prices) are often endogenous, making conventional quantile regression inconsistent and hence inappropriate for recovering the causal effects of these variables on the quantiles of economic outcomes. In order to address this problem, we develop a model of quantile treatment effects (QTE) in the presence of endogeneity and obtain conditions for identification of the QTE without functional form assumptions. The principal feature of the model is the imposition of conditions that restrict the evolution of ranks across treatment states. This feature allows us to overcome the endogeneity problem and recover the true QTE through the use of instrumental variables. The proposed model can also be equivalently viewed as a structural simultaneous equation model with nonadditive errors, where QTE can be interpreted as the structural quantile effects (SQE).

Inference for Dependent Data with Learned Clusters

The Review of Economics and Statistics 2025 107(6), 1684-1701 open access
This article presents and analyzes an approach to cluster-based inference for dependent data. The primary setting considered here is with spatially indexed data in which the dependence structure of observed random variables is characterized by a known, observed dissimilarity measure over spatial indices. Observations are partitioned into clusters with the use of an unsupervised clustering algorithm applied to the dissimilarity measure. Once the partition into clusters is learned, a cluster-based inference procedure is applied to a statistical hypothesis testing procedure. The procedure proposed in the article allows the number of clusters to depend on the data, which gives researchers a principled method for choosing an appropriate clustering level. The article gives conditions under which the proposed procedure asymptotically attains correct size. A simulation study shows that the proposed procedure attains near nominal size in finite samples in a variety of statistical testing problems with dependent data.

The Effects of 401(K) Participation on the Wealth Distribution: An Instrumental Quantile Regression Analysis

The Review of Economics and Statistics 2004 86(3), 735-751
We use instrumental quantile regression approach to examine the effects of 401(k) plans on wealth using data from the Survey of Income and Program Participation. Using 401(k) eligibility as an instrument for 401(k) participation, we estimate the quantile treatment effects of participation in a 401(k) plan on several measures of wealth. The results show the effects of 401(k) participation on net financial assets are positive and significant over the entire range of the asset distribution, and that the increase in the low tail of the assets distribution appears to translate completely into an increase in wealth. However, there is significant evidence of substitution from other forms of wealth in the upper tail of the distribution.

Targeted Undersmoothing: Sensitivity Analysis for Sparse Estimators

The Review of Economics and Statistics 2023 105(1), 101-112 open access
This paper proposes a procedure for assessing the sensitivity of inferential conclusions for functionals of sparse high-dimensional models following model selection. The proposed procedure is called targeted undersmoothing. Functionals considered include dense functionals that may depend on many or all elements of the high-dimensional parameter vector. The sensitivity analysis is based on systematic enlargements of an initially selected model. By varying the enlargements, one can conduct sensitivity analysis about the strength of empirical conclusions to model selection mistakes. We illustrate the procedure's performance through simulation experiments and two empirical examples.

Plausibly Exogenous

The Review of Economics and Statistics 2012 94(1), 260-272
Instrumental variable (IV) methods are widely used to identify causal effects in models with endogenous explanatory variables. Often the instrument exclusion restriction that underlies the validity of the usual IV inference is suspect; that is, instruments are only plausibly exogenous. We present practical methods for performing inference while relaxing the exclusion restriction. We illustrate the approaches with empirical examples that examine the effect of 401(k) participation on asset accumulation, price elasticity of demand for margarine, and returns to schooling. We find that inference is informative even with a substantial relaxation of the exclusion restriction in two of the three cases.

Pre-Event Trends in the Panel Event-Study Design

American Economic Review 2019 109(9), 3307-3338 open access
We consider a linear panel event-study design in which unobserved confounds may be related both to the outcome and to the policy variable of interest. We provide sufficient conditions to identify the causal effect of the policy by exploiting covariates related to the policy only through the confounds. Our model implies a set of moment equations that are linear in parameters. The effect of the policy can be estimated by 2SLS, and causal inference is valid even when endogeneity leads to pre-event trends (“pre-trends”) in the outcome. Alternative approaches perform poorly in our simulations.

Post-Selection and Post-Regularization Inference in Linear Models with Many Controls and Instruments

American Economic Review 2015 105(5), 486-490 open access
We consider estimation of and inference about coefficients on endogenous variables in a linear instrumental variables model where the number of instruments and exogenous control variables are each allowed to be larger than the sample size. We work within an approximately sparse framework that maintains that the signal available in the instruments and control variables may be effectively captured by a small number of the available variables. We provide a LASSO-based method for this setting which provides uniformly valid inference about the coefficients on endogenous variables. We illustrate the method through an application to demand estimation.