IDEAS home Printed from https://ideas.repec.org/p/ifs/cemmap/47-16.html
   My bibliography  Save this paper

On cross-validated Lasso

Author

Listed:
  • Denis Chetverikov

    (Institute for Fiscal Studies and UCLA)

  • . .

    (Institute for Fiscal Studies)

Abstract

In this paper, we derive a rate of convergence of the Lasso estimator when the penalty parameter ? for the estimator is chosen using K-fold cross-validation; in particular, we show that in the model with Gaussian noise and under fairly general assumptions on the candidate set of values of ?, the prediction norm of the estimation error of the cross-validated Lasso estimator is with high probability bounded from above up-to a constant by (s log p/n)1/2 (log7/8n) as long as p log n/n = o(1) and some other mild regularity conditions are satisfi ed where n is the sample size of available data, p is the number of covariates, and s is the number of non-zero coefficients in the model. Thus, the cross-validated Lasso estimator achieves the fastest possible rate of convergence up-to the logarithmic factor log7/8 n. In addition, we derive a sparsity bound for the cross-validated Lasso estimator; in particular, we show that under the same conditions as above, the number of non-zero coefficients of the estimator is with high probability bounded from above up-to a constant by s log5 n. Finally, we show that our proof technique generates non-trivial bounds on the prediction norm of the estimation error of the cross-validated Lasso estimator even if p is much larger than n and the assumption of Gaussian noise fails; in particular, the prediction norm of the estimation error is with high-probability bounded from above up-to a constant by (s log2(pn) / n)1/4 under mild regularity conditions.

Suggested Citation

  • Denis Chetverikov & . ., 2016. "On cross-validated Lasso," CeMMAP working papers CWP47/16, Centre for Microdata Methods and Practice, Institute for Fiscal Studies.
  • Handle: RePEc:ifs:cemmap:47/16
    as

    Download full text from publisher

    File URL: https://www.ifs.org.uk/uploads/cemmap/wps/cwp471616.pdf
    Download Restriction: no
    ---><---

    References listed on IDEAS

    as
    1. Albert Saiz & Uri Simonsohn, 2013. "Proxying For Unobservable Variables With Internet Document-Frequency," Journal of the European Economic Association, European Economic Association, vol. 11(1), pages 137-165, February.
    2. Alexandre Belloni & Victor Chernozhukov & Ivan Fernandez-Val & Christian Hansen, 2013. "Program evaluation with high-dimensional data," CeMMAP working papers CWP77/13, Centre for Microdata Methods and Practice, Institute for Fiscal Studies.
    3. Alexandre Belloni & Victor Chernozhukov, 2011. "High Dimensional Sparse Econometric Models: An Introduction," Papers 1106.5242, arXiv.org, revised Sep 2011.
    4. Victor Chernozhukov & Denis Chetverikov & Kengo Kato, 2014. "Central limit theorems and bootstrap in high dimensions," CeMMAP working papers 49/14, Institute for Fiscal Studies.
    5. Denis Chetverikov & Daniel Wilhelm, 2017. "Nonparametric Instrumental Variable Estimation Under Monotonicity," Econometrica, Econometric Society, vol. 85, pages 1303-1320, July.
    6. Victor Chernozhukov & Denis Chetverikov & Kengo Kato, 2012. "Gaussian approximations and multiplier bootstrap for maxima of sums of high-dimensional random vectors," Papers 1212.6906, arXiv.org, revised Jan 2018.
    7. Belloni, Alexandre & Chernozhukov, Victor & Chetverikov, Denis & Kato, Kengo, 2015. "Some new asymptotic theory for least squares series: Pointwise and uniform results," Journal of Econometrics, Elsevier, vol. 186(2), pages 345-366.
    8. David Cesarini & Christopher T. Dawes & Magnus Johannesson & Paul Lichtenstein & Björn Wallace, 2009. "Genetic Variation in Preferences for Giving and Risk Taking," The Quarterly Journal of Economics, President and Fellows of Harvard College, vol. 124(2), pages 809-842.
    Full references (including those not matched with items on IDEAS)

    Citations

    Citations are extracted by the CitEc Project, subscribe to its RSS feed for this item.
    as


    Cited by:

    1. Alain Hecq & Luca Margaritella & Stephan Smeekes, 2023. "Granger Causality Testing in High-Dimensional VARs: A Post-Double-Selection Procedure," Journal of Financial Econometrics, Oxford University Press, vol. 21(3), pages 915-958.
    2. Denis Chetverikov & Jesper R.-V. Sørensen, 2021. "Analytic and Bootstrap-after-Cross-Validation Methods for Selecting Penalty Parameters of High-Dimensional M-Estimators," Discussion Papers 21-04, University of Copenhagen. Department of Economics.
    3. Damian Kozbur, 2017. "Testing-Based Forward Model Selection," American Economic Review, American Economic Association, vol. 107(5), pages 266-269, May.
    4. Michael C. Knaus, 2021. "A double machine learning approach to estimate the effects of musical practice on student’s skills," Journal of the Royal Statistical Society Series A, Royal Statistical Society, vol. 184(1), pages 282-300, January.
    5. Denis Chetverikov & Jesper Riis-Vestergaard S{o}rensen, 2021. "Selecting Penalty Parameters of High-Dimensional M-Estimators using Bootstrapping after Cross-Validation," Papers 2104.04716, arXiv.org, revised Aug 2023.
    6. Smeekes, Stephan & Wijler, Etienne, 2021. "An automated approach towards sparse single-equation cointegration modelling," Journal of Econometrics, Elsevier, vol. 221(1), pages 247-276.

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Denis Chetverikov & . ., 2016. "On cross-validated Lasso," CeMMAP working papers 47/16, Institute for Fiscal Studies.
    2. Victor Chernozhukov & Denis Chetverikov & Kengo Kato, 2013. "Testing Many Moment Inequalities," CeMMAP working papers 65/13, Institute for Fiscal Studies.
    3. Hansen, Christian & Liao, Yuan, 2019. "The Factor-Lasso And K-Step Bootstrap Approach For Inference In High-Dimensional Economic Applications," Econometric Theory, Cambridge University Press, vol. 35(3), pages 465-509, June.
    4. Alexandre Belloni & Victor Chernozhukov & Abhishek Kaul, 2017. "Confidence bands for coefficients in high dimensional linear models with error-in-variables," CeMMAP working papers CWP22/17, Centre for Microdata Methods and Practice, Institute for Fiscal Studies.
    5. Victor Chernozhukov & Denis Chetverikov & Mert Demirer & Esther Duflo & Christian Hansen & Whitney K. Newey, 2016. "Double machine learning for treatment and causal parameters," CeMMAP working papers 49/16, Institute for Fiscal Studies.
    6. Damian Kozbur, 2013. "Inference in additively separable models with a high-dimensional set of conditioning variables," ECON - Working Papers 284, Department of Economics - University of Zurich, revised Apr 2018.
    7. Victor Chernozhukov & Vira Semenova, 2018. "Simultaneous inference for Best Linear Predictor of the Conditional Average Treatment Effect and other structural functions," CeMMAP working papers CWP40/18, Centre for Microdata Methods and Practice, Institute for Fiscal Studies.
    8. Chetverikov, Denis & Wilhelm, Daniel & Kim, Dongwoo, 2021. "An Adaptive Test Of Stochastic Monotonicity," Econometric Theory, Cambridge University Press, vol. 37(3), pages 495-536, June.
    9. Peter Horvath & Jia Li & Zhipeng Liao & Andrew J. Patton, 2022. "A consistent specification test for dynamic quantile models," Quantitative Economics, Econometric Society, vol. 13(1), pages 125-151, January.
    10. Alexandre Belloni & Victor Chernozhukov & Lie Wang, 2013. "Pivotal estimation via square-root lasso in nonparametric regression," CeMMAP working papers CWP62/13, Centre for Microdata Methods and Practice, Institute for Fiscal Studies.
    11. Achim Ahrens & Christian B. Hansen & Mark E. Schaffer, 2020. "lassopack: Model selection and prediction with regularized regression in Stata," Stata Journal, StataCorp LP, vol. 20(1), pages 176-235, March.
    12. Byunghoon Kang, 2018. "Inference in Nonparametric Series Estimation with Specification Searches for the Number of Series Terms," Working Papers 240829404, Lancaster University Management School, Economics Department.
    13. Li, Jia & Liao, Zhipeng, 2020. "Uniform nonparametric inference for time series," Journal of Econometrics, Elsevier, vol. 219(1), pages 38-51.
    14. Chernozhukov, Victor & Chetverikov, Denis & Kato, Kengo, 2016. "Empirical and multiplier bootstraps for suprema of empirical processes of increasing complexity, and related Gaussian couplings," Stochastic Processes and their Applications, Elsevier, vol. 126(12), pages 3632-3651.
    15. Max H. Farrell, 2013. "Robust Inference on Average Treatment Effects with Possibly More Covariates than Observations," Papers 1309.4686, arXiv.org, revised Feb 2018.
    16. Victor Chernozhukov & Whitney K. Newey & Andres Santos, 2023. "Constrained Conditional Moment Restriction Models," Econometrica, Econometric Society, vol. 91(2), pages 709-736, March.
    17. Matias D. Cattaneo & Ricardo P. Masini & William G. Underwood, 2022. "Yurinskii's Coupling for Martingales," Papers 2210.00362, arXiv.org, revised Mar 2024.
    18. Farrell, Max H., 2015. "Robust inference on average treatment effects with possibly more covariates than observations," Journal of Econometrics, Elsevier, vol. 189(1), pages 1-23.
    19. Ruben Dezeure & Peter Bühlmann & Cun-Hui Zhang, 2017. "High-dimensional simultaneous inference with the bootstrap," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 26(4), pages 685-719, December.
    20. Christian Hansen & Damian Kozbur & Sanjog Misra, 2016. "Targeted undersmoothing," ECON - Working Papers 282, Department of Economics - University of Zurich, revised Apr 2018.

    More about this item

    Keywords

    Cross-Validated Lasso;

    NEP fields

    This paper has been announced in the following NEP Reports:

    Statistics

    Access and download statistics

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:ifs:cemmap:47/16. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Emma Hyman (email available below). General contact details of provider: https://edirc.repec.org/data/cmifsuk.html .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.