Advanced Search
MyIDEAS: Login

Large covariance estimation by thresholding principal orthogonal complements

Contents:

Author Info

  • Jianqing Fan
  • Yuan Liao
  • Martina Mincheva

Abstract

This paper deals with estimation of high-dimensional covariance with a conditional sparsity structure, which is the composition of a low-rank matrix plus a sparse matrix. By assuming sparse error covariance matrix in a multi-factor model, we allow the presence of the cross-sectional correlation even after taking out common but unobservable factors. We introduce the Principal Orthogonal complEment Thresholding (POET) method to explore such an approximate factor structure. The POET estimator includes the sample covariance matrix, the factor-based covariance matrix (Fan, Fan and Lv, 2008), the thresholding estimator (Bickel and Levina, 2008) and the adaptive thresholding estimator (Cai and Liu, 2011) as specic examples. We provide mathematical insights when the factor analysis is approximately the same as the principal component analysis for high dimensional data. The rates of convergence of the sparse residual covariance matrix and the conditional sparse covariance matrix are studied under various norms, including the spectral norm. It is shown that the impact of estimating the unknown factors vanishes as the dimensionality increases. The uniform rates of convergence for the unobserved factors and their factor loadings are derived. The asymptotic results are also veried by extensive simulation studies.

(This abstract was borrowed from another version of this item.)

Download Info

If you experience problems downloading a file, check if you have the proper application to view it first. In case of further problems read the IDEAS help page. Note that these files are not on the IDEAS site. Please be patient as the files may be large.
File URL: http://hdl.handle.net/10.1111/rssb.12016
Download Restriction: Access to full text is restricted to subscribers.

As the access to this document is restricted, you may want to look for a different version under "Related research" (further below) or search for a different version of it.

Bibliographic Info

Article provided by Royal Statistical Society in its journal Journal of the Royal Statistical Society: Series B (Statistical Methodology).

Volume (Year): 75 (2013)
Issue (Month): 4 (09)
Pages: 603-680

as in new window
Handle: RePEc:bla:jorssb:v:75:y:2013:i:4:p:603-680

Contact details of provider:
Postal: 12 Errol Street, London EC1Y 8LX, United Kingdom
Phone: -44-171-638-8998
Fax: -44-171-256-7598
Email:
Web page: http://wileyonlinelibrary.com/journal/rssb
More information through EDIRC

Order Information:
Web: http://ordering.onlinelibrary.wiley.com/subs.asp?ref=1467-9868&doi=10.1111/(ISSN)1467-9868

Related research

Keywords:

Other versions of this item:

Find related papers by JEL classification:

References

References listed on IDEAS
Please report citation or reference errors to , or , if you are the registered author of the cited work, log in to your RePEc Author Service profile, click on "citations" and make appropriate adjustments.:
as in new window
  1. Ross, Stephen A., 1976. "The arbitrage theory of capital asset pricing," Journal of Economic Theory, Elsevier, vol. 13(3), pages 341-360, December.
  2. Cai, Tony & Liu, Weidong, 2011. "Adaptive Thresholding for Sparse Covariance Matrix Estimation," Journal of the American Statistical Association, American Statistical Association, vol. 106(494), pages 672-684.
  3. Fan, Jianqing & Fan, Yingying & Lv, Jinchi, 2008. "High dimensional covariance matrix estimation using a factor model," Journal of Econometrics, Elsevier, vol. 147(1), pages 186-197, November.
  4. Shen, Haipeng & Huang, Jianhua Z., 2008. "Sparse principal component analysis via regularized low rank matrix approximation," Journal of Multivariate Analysis, Elsevier, vol. 99(6), pages 1015-1034, July.
  5. Efron, Bradley, 2010. "Correlated z-Values and the Accuracy of Large-Scale Statistical Estimates," Journal of the American Statistical Association, American Statistical Association, vol. 105(491), pages 1042-1055.
  6. Catherine Doz & Domenico Giannone & Lucrezia Reichlin, 2006. "A Two-step estimator for large approximate dynamic factor models based on Kalman filtering," THEMA Working Papers 2006-23, THEMA (THéorie Economique, Modélisation et Applications), Université de Cergy-Pontoise.
  7. Hallin, Marc & Liska, Roman, 2007. "Determining the Number of Factors in the General Dynamic Factor Model," Journal of the American Statistical Association, American Statistical Association, vol. 102, pages 603-617, June.
  8. Chamberlain, Gary & Rothschild, Michael, 1983. "Arbitrage, Factor Structure, and Mean-Variance Analysis on Large Asset Markets," Econometrica, Econometric Society, vol. 51(5), pages 1281-304, September.
  9. Fama, Eugene F & French, Kenneth R, 1992. " The Cross-Section of Expected Stock Returns," Journal of Finance, American Finance Association, vol. 47(2), pages 427-65, June.
  10. Jushan Bai, 2003. "Inferential Theory for Factor Models of Large Dimensions," Econometrica, Econometric Society, vol. 71(1), pages 135-171, January.
  11. Rothman, Adam J. & Levina, Elizaveta & Zhu, Ji, 2009. "Generalized Thresholding of Large Covariance Matrices," Journal of the American Statistical Association, American Statistical Association, vol. 104(485), pages 177-186.
  12. Alexei Onatski, 2010. "Determining the Number of Factors from Empirical Distribution of Eigenvalues," The Review of Economics and Statistics, MIT Press, vol. 92(4), pages 1004-1016, November.
  13. Mario Forni & Marc Hallin & Lucrezia Reichlin & Marco Lippi, 2000. "The generalised dynamic factor model: identification and estimation," ULB Institutional Repository 2013/10143, ULB -- Universite Libre de Bruxelles.
  14. Efron, Bradley, 2007. "Correlation and Large-Scale Simultaneous Significance Testing," Journal of the American Statistical Association, American Statistical Association, vol. 102, pages 93-103, March.
  15. M. Hashem Pesaran, 2004. "Estimation and Inference in Large Heterogeneous Panels with a Multifactor Error Structure," CESifo Working Paper Series 1331, CESifo Group Munich.
  16. Kapetanios, George, 2010. "A Testing Procedure for Determining the Number of Factors in Approximate Factor Models With Large Datasets," Journal of Business & Economic Statistics, American Statistical Association, vol. 28(3), pages 397-409.
  17. Stock J.H. & Watson M.W., 2002. "Forecasting Using Principal Components From a Large Number of Predictors," Journal of the American Statistical Association, American Statistical Association, vol. 97, pages 1167-1179, December.
  18. Johnstone, Iain M. & Lu, Arthur Yu, 2009. "On Consistency and Sparsity for Principal Components Analysis in High Dimensions," Journal of the American Statistical Association, American Statistical Association, vol. 104(486), pages 682-693.
  19. Xi Luo, 2011. "Recovering Model Structures from Large Low Rank and Sparse Covariance Matrix Estimation," Papers 1111.1133, arXiv.org, revised Mar 2013.
  20. Jianqing Fan & Jingjin Zhang & Ke Yu, 2008. "Asset Allocation and Risk Assessment with Gross Exposure Constraints for Vast Portfolios," Papers 0812.2604, arXiv.org.
  21. Ahn, Seung Chan & Hoon Lee, Young & Schmidt, Peter, 2001. "GMM estimation of linear panel data models with time-varying individual effects," Journal of Econometrics, Elsevier, vol. 101(2), pages 219-255, April.
Full references (including those not matched with items on IDEAS)

Citations

Citations are extracted by the CitEc Project, subscribe to its RSS feed for this item.
as in new window

Cited by:
  1. Jianqing Fan & Yuan Liao & Xiaofeng Shi, 2013. "Risks of Large Portfolios," Papers 1302.0926, arXiv.org.
  2. Matteo Barigozzi & Christian T. Brownlees, 2013. "Nets: Network estimation for time series," Economics Working Papers 1391, Department of Economics and Business, Universitat Pompeu Fabra.
  3. Taras Bodnar & Nestor Parolya & Wolfgang Schmid, 2014. "Estimation of the Global Minimum Variance Portfolio in High Dimensions," Papers 1406.0437, arXiv.org.
  4. Bai, Jushan & Liao, Yuan, 2012. "Efficient Estimation of Approximate Factor Models," MPRA Paper 41558, University Library of Munich, Germany.
  5. Natalia Bailey & M. Hashem Pesaran & L. Vanessa Smith, 2014. "A Multiple Testing Approach to the Regularisation of Large Sample Correlation Matrices," CESifo Working Paper Series 4834, CESifo Group Munich.

Lists

This item is not listed on Wikipedia, on a reading list or among the top items on IDEAS.

Statistics

Access and download statistics

Corrections

When requesting a correction, please mention this item's handle: RePEc:bla:jorssb:v:75:y:2013:i:4:p:603-680. See general information about how to correct material in RePEc.

For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: (Wiley-Blackwell Digital Licensing) or (Christopher F. Baum).

If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

If references are entirely missing, you can add them using this form.

If the full references list an item that is present in RePEc, but the system did not link to it, you can help with this form.

If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your profile, as there may be some citations waiting for confirmation.

Please note that corrections may take a couple of weeks to filter through the various RePEc services.