A Nonparametric Test for Equality of Distributions with Mixed Categorical and Continuous Data
In this paper we consider the problem of testing for equality of two density or two conditional density functions defined over mixed discrete and continuous variables. We smooth both the discrete and continuous variables, with the smoothing parameters chosen via least-squares cross-validation. The test statistics are shown to have (asymptotic) normal null distributions. However, we advocate the use of bootstrap methods in order to better approximate their null distribution in nite-sample settings. Simulations show that the proposed tests have better power than both conventional frequency-based tests and smoothing tests based on ad hoc smoothing parameter selection, while a demonstrative empirical application to the joint distribution of earnings and educational attainment underscores the utility of the proposed approach in mixed data settings.
|Date of creation:||Oct 2008|
|Date of revision:|
|Contact details of provider:|| Web page: http://economics.emory.edu/home/journals/|
More information through EDIRC
References listed on IDEAS
Please report citation or reference errors to , or , if you are the registered author of the cited work, log in to your RePEc Author Service profile, click on "citations" and make appropriate adjustments.:
- Russell Davidson & James G. MacKinnon, 2001.
"Bootstrap Tests: How Many Bootstraps?,"
1036, Queen's University, Department of Economics.
- Hall, Peter, 1984. "Central limit theorem for integrated square error of multivariate nonparametric density estimators," Journal of Multivariate Analysis, Elsevier, vol. 14(1), pages 1-16, February.
- Ahmad, Ibrahim A. & Li, Qi, 1997. "Testing independence by nonparametric kernel method," Statistics & Probability Letters, Elsevier, vol. 34(2), pages 201-210, June.
- Kiefer, Nicholas M. & Racine, Jeffrey S., 2008. "The Smooth Colonel Meets the Reverend," Working Papers 08-01, Cornell University, Center for Analytic Economics.
- Fan, Yanqin & Li, Qi, 2000. "Consistent Model Specification Tests," Econometric Theory, Cambridge University Press, vol. 16(06), pages 1016-1041, December.
- Robinson, P M, 1991. "Consistent Nonparametric Entropy-Based Testing," Review of Economic Studies, Wiley Blackwell, vol. 58(3), pages 437-53, May.
- Racine, Jeffrey S. & Maasoumi, Esfandiar, 2007. "A versatile and robust metric entropy test of time-reversibility, and other hypotheses," Journal of Econometrics, Elsevier, vol. 138(2), pages 547-567, June.
- Li, Qi & Racine, Jeff, 2003. "Nonparametric estimation of distributions with categorical and continuous data," Journal of Multivariate Analysis, Elsevier, vol. 86(2), pages 266-292, August.
- Grund, B. & Hall, P., 1993. "On the Performance of Kernel Estimators for High-Dimensional, Sparse Binary Data," Journal of Multivariate Analysis, Elsevier, vol. 44(2), pages 321-344, February.
- Peter Hall & Jeff Racine & Qi Li, 2004. "Cross-Validation and the Estimation of Conditional Probability Densities," Journal of the American Statistical Association, American Statistical Association, vol. 99, pages 1015-1026, December.
- Gordon Anderson, 2001. "The Power And Size Of Nonparametric Tests For Common Distributional Characteristics," Econometric Reviews, Taylor & Francis Journals, vol. 20(1), pages 1-30.
- Fan, Yanqin, 1998. "Goodness-Of-Fit Tests Based On Kernel Density Estimators With Fixed Smoothing Parameters," Econometric Theory, Cambridge University Press, vol. 14(05), pages 604-621, October.
- Yongmiao Hong & Halbert White, 2005. "Asymptotic Distribution Theory for Nonparametric Entropy Measures of Serial Dependence," Econometrica, Econometric Society, vol. 73(3), pages 837-901, 05.
- Peter Hall & Qi Li & Jeffrey S. Racine, 2007. "Nonparametric Estimation of Regression Functions in the Presence of Irrelevant Regressors," The Review of Economics and Statistics, MIT Press, vol. 89(4), pages 784-789, November.
- Anderson, N. H. & Hall, P. & Titterington, D. M., 1994. "Two-Sample Test Statistics for Measuring Discrepancies Between Two Multivariate Probability Density Functions Using Kernel-Based Density Estimates," Journal of Multivariate Analysis, Elsevier, vol. 50(1), pages 41-54, July.
When requesting a correction, please mention this item's handle: RePEc:emo:wp2003:0805. See general information about how to correct material in RePEc.
For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: (Sue Mialon)
If references are entirely missing, you can add them using this form.