EconBase
← All papers

Influence Analysis with Panel Data

Annalivia Polselli

arXiv 9 Dec 2023 · Econometrics

arXiv:2312.05700 · PDF · DOI · OpenAlex · Extracted main text

Abstract

The presence of units with extreme values in the dependent and/or independent variables (i.e., vertical outliers, leveraged data) has the potential to severely bias regression coefficients and/or standard errors. This is common with short panel data because the researcher cannot advocate asymptotic theory. Example include cross-country studies, cell-group analyses, and field or laboratory experimental studies, where the researcher is forced to use few cross-sectional observations repeated over time due to the structure of the data or research design. Available diagnostic tools may fail to properly detect these anomalies, because they are not designed for panel data. In this paper, we formalise statistical measures for panel data models with fixed effects to quantify the degree of leverage and outlyingness of units, and the joint and conditional influences of pairs of units. We first develop a method to visually detect anomalous units in a panel data set, and identify their type. Second, we investigate the effect of these units on LS estimates, and on other units' influence on the estimated parameters. To illustrate and validate the proposed method, we use a synthetic data set contaminated with different types of anomalous units. We also provide an empirical example.

Citation extraction

31
references
98
in-text mentions
31
distinct cited
1
self-citations
6,780
main-text words

appendix boundary found by appendix_command · 62% of the source is main text. Read the extracted text to check this.

Most heavily cited references

The works this paper leans on most, across its whole bibliography — not restricted to papers in our corpus. Ranked by composite intensity, which combines how often a work is mentioned, how many sections mention it, and how much of that falls in the main text rather than the appendix.

ReferenceIntensityMentionsSectionsMain text
1Belotti, F. and Peracchi, F (2020) Fast leave-one-out methods for inference, model selection, and diagnostic checking1.00064100%
2MacKinnon, J. G., Nielsen, M. ., and Webb, M. D (2023) Leverage, influence, and the jackknife in clustered regression models: Reliable inference using summclust1.00064100%
3Verardi, V. and Croux, C (2009) Robust regression in stata1.00064100%
4Bramati, M. C. and Croux, C (2007) Robust estimators for the fixed effects panel data model0.96510590%
5Chesher, A. and Jewitt, I (1987) The bias of a heteroskedasticity consistent covariance matrix estimator0.92844100%
6MacKinnon, J. G. and White, H (1985) Some heteroskedasticity-consistent covariance matrix estimators with improved finite sample properties0.92844100%
7MacKinnon, J. G., Nielsen, M. ., and Webb, M. D (2023) Cluster-robust inference: A guide to empirical practice0.92843100%
8Rousseeuw, P. J (1991) A diagnostic plot for regression outliers and leverage points0.92843100%
9Berka, M., Devereux, M. B., and Engel, C (2018) Real exchange rates and sectoral productivity in the eurozone0.87462100%
10Jiao, X (2022) A simple robust procedure in instrumental variables regression0.84333100%

Showing the top 10 of 31 scored citations.