Cross-validation Bandwidth Matrices for Multivariate Kernel Density Estimation
AbstractThe performance of multivariate kernel density estimates depends crucially on the choice of bandwidth matrix, but progress towards developing good bandwidth matrix selectors has been relatively slow. In particular, previous studies of cross-validation (CV) methods have been restricted to biased and unbiased CV selection of diagonal bandwidth matrices. However, for certain types of target density the use of full (i.e. unconstrained) bandwidth matrices offers the potential for significantly improved density estimation. In this paper, we generalize earlier work from diagonal to full bandwidth matrices, and develop a smooth cross-validation (SCV) methodology for multivariate data. We consider optimization of the SCV technique with respect to a pilot bandwidth matrix. All the CV methods are studied using asymptotic analysis, simulation experiments and real data analysis. The results suggest that SCV for full bandwidth matrices is the most reliable of the CV methods. We also observe that experience from the univariate setting can sometimes be a misleading guide for understanding bandwidth selection in the multivariate case. Copyright 2005 Board of the Foundation of the Scandinavian Journal of Statistics..
Download InfoIf you experience problems downloading a file, check if you have the proper application to view it first. In case of further problems read the IDEAS help page. Note that these files are not on the IDEAS site. Please be patient as the files may be large.
As the access to this document is restricted, you may want to look for a different version under "Related research" (further below) or search for a different version of it.
Bibliographic InfoArticle provided by Danish Society for Theoretical Statistics & Finnish Statistical Society & Norwegian Statistical Association & Swedish Statistical Association in its journal Scandinavian Journal of Statistics.
Volume (Year): 32 (2005)
Issue (Month): 3 ()
Contact details of provider:
Web page: http://www.blackwellpublishing.com/journal.asp?ref=0303-6898
You can help add them by filling out this form.
CitEc Project, subscribe to its RSS feed for this item.
- J. Chacón & T. Duong, 2010. "Multivariate plug-in bandwidth selection with unconstrained pilot bandwidth matrices," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer, vol. 19(2), pages 375-398, August.
- Shuowen Hu & D.S. Poskitt & Xibin Zhang, 2010.
"Bayesian Adaptive Bandwidth Kernel Density Estimation of Irregular Multivariate Distributions,"
Monash Econometrics and Business Statistics Working Papers
21/10, Monash University, Department of Econometrics and Business Statistics.
- Hu, Shuowen & Poskitt, D.S. & Zhang, Xibin, 2012. "Bayesian adaptive bandwidth kernel density estimation of irregular multivariate distributions," Computational Statistics & Data Analysis, Elsevier, vol. 56(3), pages 732-740.
- Billings, Stephen B. & Johnson, Erik B., 2012. "A non-parametric test for industrial specialization," Journal of Urban Economics, Elsevier, vol. 71(3), pages 312-331.
- Nicolai, R.P. & Koning, A.J., 2006. "A general framework for statistical inference on discrete event systems," Econometric Institute Research Papers EI 2006-45, Erasmus University Rotterdam, Erasmus School of Economics (ESE), Econometric Institute.
- Duong, Tarn & Cowling, Arianna & Koch, Inge & Wand, M.P., 2008. "Feature significance for multivariate kernel density estimation," Computational Statistics & Data Analysis, Elsevier, vol. 52(9), pages 4225-4242, May.
- Horová, Ivana & Koláček, Jan & Vopatová, Kamila, 2013. "Full bandwidth matrix selectors for gradient kernel density estimate," Computational Statistics & Data Analysis, Elsevier, vol. 57(1), pages 364-376.
- Rob J. Hyndman & Han Lin Shang, 2008. "Rainbow plots, Bagplots and Boxplots for Functional Data," Monash Econometrics and Business Statistics Working Papers 9/08, Monash University, Department of Econometrics and Business Statistics.
- Schoch, Tobias & Staub, Kaspar & Pfister, Christian, 2012. "Social inequality and the biological standard of living: An anthropometric analysis of Swiss conscription data, 1875–1950," Economics & Human Biology, Elsevier, vol. 10(2), pages 154-173.
- Filippone, Maurizio & Sanguinetti, Guido, 2011. "Approximate inference of the bandwidth in multivariate kernel density estimation," Computational Statistics & Data Analysis, Elsevier, vol. 55(12), pages 3104-3122, December.
- Seppo Pulkkinen & Marko Mäkelä & Napsu Karmitsa, 2013. "A continuation approach to mode-finding of multivariate Gaussian mixtures and kernel density estimates," Journal of Global Optimization, Springer, vol. 56(2), pages 459-487, June.
For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: (Wiley-Blackwell Digital Licensing) or (Christopher F. Baum).
If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.
If references are entirely missing, you can add them using this form.
If the full references list an item that is present in RePEc, but the system did not link to it, you can help with this form.
If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your profile, as there may be some citations waiting for confirmation.
Please note that corrections may take a couple of weeks to filter through the various RePEc services.