IDEAS home Printed from https://ideas.repec.org/a/eee/stapro/v181y2022ics0167715221002376.html
   My bibliography  Save this article

Reproducible feature selection in high-dimensional accelerated failure time models

Author

Listed:
  • Dong, Yan
  • Li, Daoji
  • Zheng, Zemin
  • Zhou, Jia

Abstract

We propose a new feature selection procedure with guaranteed FDR control for high-dimensional AFT models, which is among the first attempts of reproducible learning in survival analysis. The effectiveness of the proposed method is theoretically and numerically demonstrated.

Suggested Citation

  • Dong, Yan & Li, Daoji & Zheng, Zemin & Zhou, Jia, 2022. "Reproducible feature selection in high-dimensional accelerated failure time models," Statistics & Probability Letters, Elsevier, vol. 181(C).
  • Handle: RePEc:eee:stapro:v:181:y:2022:i:c:s0167715221002376
    DOI: 10.1016/j.spl.2021.109275
    as

    Download full text from publisher

    File URL: http://www.sciencedirect.com/science/article/pii/S0167715221002376
    Download Restriction: Full text for ScienceDirect subscribers only

    File URL: https://libkey.io/10.1016/j.spl.2021.109275?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    As the access to this document is restricted, you may want to search for a different version of it.

    References listed on IDEAS

    as
    1. Stute, W., 1993. "Consistent Estimation Under Random Censorship When Covariables Are Present," Journal of Multivariate Analysis, Elsevier, vol. 45(1), pages 89-103, April.
    2. Jian Huang & Shuangge Ma & Huiliang Xie, 2006. "Regularized Estimation in the Accelerated Failure Time Model with High-Dimensional Covariates," Biometrics, The International Biometric Society, vol. 62(3), pages 813-820, September.
    3. Yingying Fan & Emre Demirkaya & Gaorong Li & Jinchi Lv, 2020. "RANK: Large-Scale Inference With Graphical Nonlinear Knockoffs," Journal of the American Statistical Association, Taylor & Francis Journals, vol. 115(529), pages 362-379, January.
    4. Anders Gorst-Rasmussen & Thomas Scheike, 2013. "Independent screening for single-index hazard rate models with ultrahigh dimensional features," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 75(2), pages 217-246, March.
    5. Emmanuel Candès & Yingying Fan & Lucas Janson & Jinchi Lv, 2018. "Panning for gold: ‘model‐X’ knockoffs for high dimensional controlled variable selection," Journal of the Royal Statistical Society Series B, Royal Statistical Society, vol. 80(3), pages 551-577, June.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Pan, Yingli, 2022. "Feature screening and FDR control with knockoff features for ultrahigh-dimensional right-censored data," Computational Statistics & Data Analysis, Elsevier, vol. 173(C).
    2. Ruoqing Zhu & Ying-Qi Zhao & Guanhua Chen & Shuangge Ma & Hongyu Zhao, 2017. "Greedy outcome weighted tree learning of optimal personalized treatment rules," Biometrics, The International Biometric Society, vol. 73(2), pages 391-400, June.
    3. Xiaochao Xia & Binyan Jiang & Jialiang Li & Wenyang Zhang, 2016. "Low-dimensional confounder adjustment and high-dimensional penalized estimation for survival analysis," Lifetime Data Analysis: An International Journal Devoted to Statistical Methods and Applications for Time-to-Event Data, Springer, vol. 22(4), pages 547-569, October.
    4. Khan Md Hasinur Rahaman & Bhadra Anamika & Howlader Tamanna, 2019. "Stability selection for lasso, ridge and elastic net implemented with AFT models," Statistical Applications in Genetics and Molecular Biology, De Gruyter, vol. 18(5), pages 1-14, October.
    5. Ma, Shuangge & Dai, Ying & Huang, Jian & Xie, Yang, 2012. "Identification of breast cancer prognosis markers via integrative analysis," Computational Statistics & Data Analysis, Elsevier, vol. 56(9), pages 2718-2728.
    6. Hu, Jianwei & Chai, Hao, 2013. "Adjusted regularized estimation in the accelerated failure time model with high dimensional covariates," Journal of Multivariate Analysis, Elsevier, vol. 122(C), pages 96-114.
    7. Zhihua Sun & Yi Liu & Kani Chen & Gang Li, 2022. "Broken adaptive ridge regression for right-censored survival data," Annals of the Institute of Statistical Mathematics, Springer;The Institute of Statistical Mathematics, vol. 74(1), pages 69-91, February.
    8. Emre Demirkaya & Yang Feng & Pallavi Basu & Jinchi Lv, 2022. "Large-scale model selection in misspecified generalized linear models [Information theory and an extension of the maximum likelihood principle]," Biometrika, Biometrika Trust, vol. 109(1), pages 123-136.
    9. Wang Zhu & Wang C.Y., 2010. "Buckley-James Boosting for Survival Analysis with High-Dimensional Biomarker Data," Statistical Applications in Genetics and Molecular Biology, De Gruyter, vol. 9(1), pages 1-33, June.
    10. Dong, Qingkai & Liu, Binxia & Zhao, Hui, 2023. "Weighted least squares model averaging for accelerated failure time models," Computational Statistics & Data Analysis, Elsevier, vol. 184(C).
    11. Zhiping Qiu & Jing Qin & Yong Zhou, 2016. "Composite Estimating Equation Method for the Accelerated Failure Time Model with Length-biased Sampling Data," Scandinavian Journal of Statistics, Danish Society for Theoretical Statistics;Finnish Statistical Society;Norwegian Statistical Association;Swedish Statistical Association, vol. 43(2), pages 396-415, June.
    12. Guo, Xu & Li, Runze & Liu, Jingyuan & Zeng, Mudong, 2023. "Statistical inference for linear mediation models with high-dimensional mediators and application to studying stock reaction to COVID-19 pandemic," Journal of Econometrics, Elsevier, vol. 235(1), pages 166-179.
    13. Guessoum Zohra & Ould-Said Elias, 2009. "On nonparametric estimation of the regression function under random censorship model," Statistics & Risk Modeling, De Gruyter, vol. 26(3), pages 159-177, April.
    14. Sungwan Bang & Soo-Heang Eo & Yong Mee Cho & Myoungshic Jhun & HyungJun Cho, 2016. "Non-crossing weighted kernel quantile regression with right censored data," Lifetime Data Analysis: An International Journal Devoted to Statistical Methods and Applications for Time-to-Event Data, Springer, vol. 22(1), pages 100-121, January.
    15. Wenceslao González Manteiga & Cédric Heuchenne & César Sánchez Sellero & Alessandro Beretta, 2020. "Goodness-of-fit tests for censored regression based on artificial data points," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 29(2), pages 599-615, June.
    16. Amorim, Ana Paula & de Uña-Álvarez, Jacobo & Meira-Machado, Luís, 2011. "Presmoothing the transition probabilities in the illness-death model," Statistics & Probability Letters, Elsevier, vol. 81(7), pages 797-806, July.
    17. Uña-Álvarez, Jacobo de & González-Manteiga, Wenceslao, 1999. "Strong consistency under proportional censorship when covariables are present," Statistics & Probability Letters, Elsevier, vol. 42(3), pages 283-292, April.
    18. Sijian Wang & Bin Nan & Ji Zhu & David G. Beer, 2008. "Doubly Penalized Buckley–James Method for Survival Data with High-Dimensional Covariates," Biometrics, The International Biometric Society, vol. 64(1), pages 132-140, March.
    19. Jacobo Uña-Álvarez & Noël Veraverbeke, 2013. "Generalized copula-graphic estimator," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 22(2), pages 343-360, June.
    20. Pedro Delicado & Daniel Peña, 2023. "Understanding complex predictive models with ghost variables," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 32(1), pages 107-145, March.

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:eee:stapro:v:181:y:2022:i:c:s0167715221002376. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Catherine Liu (email available below). General contact details of provider: http://www.elsevier.com/wps/find/journaldescription.cws_home/622892/description#description .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.