IDEAS home Printed from https://ideas.repec.org/a/plo/pone00/0301912.html
   My bibliography  Save this article

Online application for the diagnosis of atherosclerosis by six genes

Author

Listed:
  • Zunlan Zhao
  • Shouhang Chen
  • Hongzhao Wei
  • Weile Ma
  • Weili Shi
  • Yixin Si
  • Jun Wang
  • Liuyi Wang
  • Xiqing Li

Abstract

Background: Atherosclerosis (AS) is a primary contributor to cardiovascular disease, leading to significant global mortality rates. Developing effective diagnostic indicators and models for AS holds the potential to substantially reduce the fatalities and disabilities associated with cardiovascular disease. Blood sample analysis has emerged as a promising avenue for facilitating diagnosis and assessing disease prognosis. Nonetheless, it lacks an accurate model or tool for AS diagnosis. Hence, the principal objective of this study is to develop a convenient, simple, and accurate model for the early detection of AS. Methods: We downloaded the expression data of blood samples from GEO databases. By dividing the mean values of housekeeping genes (meanHGs) and applying the comBat function, we aimed to reduce the batch effect. After separating the datasets into training, evaluation, and testing sets, we applied differential expression analyses (DEA) between AS and control samples from the training dataset. Then, a gradient-boosting model was used to evaluate the importance of genes and identify the hub genes. Using different machine learning algorithms, we constructed a prediction model with the highest accuracy in the testing dataset. Finally, we make the machine learning models publicly accessible by shiny app construction. Results: Seven datasets (GSE9874, GSE12288, GSE20129, GSE23746, GSE27034, GSE90074, and GSE202625), including 403 samples with AS and 325 healthy subjects, were obtained by comprehensive searching and filtering by specific requirements. The batch effect was successfully removed by dividing the meanHGs and applying the comBat function. 331 genes were found to be related to atherosclerosis by the DEA analysis between AS and health samples. The top 6 genes with the highest importance values from the gradient boosting model were identified. Out of the seven machine learning algorithms tested, the random forest model exhibited the most impressive performance in the testing datasets, achieving an accuracy exceeding 0.8. While the batch effect reduction analysis in our study could have contributed to the increased accuracy values, our comparison results further highlight the superiority of our model over the genes provided in published studies. This underscores the effectiveness of our approach in delivering superior predictive performance. The machine-learning models were then uploaded to the Shiny app’s server, making it easy for users to distinguish AS samples from normal samples. Conclusions: A prognostic Shiny application, built upon six potential atherosclerosis-associated genes, has been developed, offering an accurate diagnosis of atherosclerosis.

Suggested Citation

  • Zunlan Zhao & Shouhang Chen & Hongzhao Wei & Weile Ma & Weili Shi & Yixin Si & Jun Wang & Liuyi Wang & Xiqing Li, 2024. "Online application for the diagnosis of atherosclerosis by six genes," PLOS ONE, Public Library of Science, vol. 19(4), pages 1-15, April.
  • Handle: RePEc:plo:pone00:0301912
    DOI: 10.1371/journal.pone.0301912
    as

    Download full text from publisher

    File URL: https://journals.plos.org/plosone/article?id=10.1371/journal.pone.0301912
    Download Restriction: no

    File URL: https://journals.plos.org/plosone/article/file?id=10.1371/journal.pone.0301912&type=printable
    Download Restriction: no

    File URL: https://libkey.io/10.1371/journal.pone.0301912?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    References listed on IDEAS

    as
    1. Chao Chen & Kay Grennan & Judith Badner & Dandan Zhang & Elliot Gershon & Li Jin & Chunyu Liu, 2011. "Removing Batch Effects in Analysis of Expression Microarray Data: An Evaluation of Six Batch Adjustment Methods," PLOS ONE, Public Library of Science, vol. 6(2), pages 1-10, February.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Xia Qing & Thompson Jeffrey A. & Koestler Devin C., 2021. "Batch effect reduction of microarray data with dependent samples using an empirical Bayes approach (BRIDGE)," Statistical Applications in Genetics and Molecular Biology, De Gruyter, vol. 20(4-6), pages 101-119, December.
    2. Wei Liu & Huaqin He & Davide Chicco, 2024. "Gene signatures for cancer research: A 25-year retrospective and future avenues," PLOS Computational Biology, Public Library of Science, vol. 20(10), pages 1-10, October.
    3. Jacopo Umberto Verga & Matthew Huff & Diarmuid Owens & Bethany J. Wolf & Gary Hardiman, 2022. "Integrated Genomic and Bioinformatics Approaches to Identify Molecular Links between Endocrine Disruptors and Adverse Outcomes," IJERPH, MDPI, vol. 19(1), pages 1-24, January.
    4. Aline Talhouk & Stefan Kommoss & Robertson Mackenzie & Martin Cheung & Samuel Leung & Derek S Chiu & Steve E Kalloger & David G Huntsman & Stephanie Chen & Maria Intermaggio & Jacek Gronwald & Fong C , 2016. "Single-Patient Molecular Testing with NanoString nCounter Data Using a Reference-Based Strategy for Batch Effect Correction," PLOS ONE, Public Library of Science, vol. 11(4), pages 1-18, April.
    5. Charlotte Soneson & Sarah Gerster & Mauro Delorenzi, 2014. "Batch Effect Confounding Leads to Strong Bias in Performance Estimates Obtained by Cross-Validation," PLOS ONE, Public Library of Science, vol. 9(6), pages 1-13, June.
    6. Sean M Gibbons & Claire Duvallet & Eric J Alm, 2018. "Correcting for batch effects in case-control microbiome studies," PLOS Computational Biology, Public Library of Science, vol. 14(4), pages 1-17, April.
    7. Davide Chicco & Giuseppe Agapito, 2022. "Nine quick tips for pathway enrichment analysis," PLOS Computational Biology, Public Library of Science, vol. 18(8), pages 1-15, August.
    8. Samir Dou & Nathalie Villa-Vialaneix & Laurence Liaubet & Yvon Billon & Mario Giorgi & Hélène Gilbert & Jean-Luc Gourdine & Juliette Riquet & David Renaudeau, 2017. "1H NMR-Based metabolomic profiling method to develop plasma biomarkers for sensitivity to chronic heat stress in growing pigs," PLOS ONE, Public Library of Science, vol. 12(11), pages 1-18, November.
    9. Christian Müller & Arne Schillert & Caroline Röthemeier & David-Alexandre Trégouët & Carole Proust & Harald Binder & Norbert Pfeiffer & Manfred Beutel & Karl J Lackner & Renate B Schnabel & Laurence T, 2016. "Removing Batch Effects from Longitudinal Gene Expression - Quantile Normalization Plus ComBat as Best Approach for Microarray Transcriptome Data," PLOS ONE, Public Library of Science, vol. 11(6), pages 1-23, June.
    10. Romain Banchereau & Alejandro Jordan-Villegas & Monica Ardura & Asuncion Mejias & Nicole Baldwin & Hui Xu & Elizabeth Saye & Jose Rossello-Urgell & Phuong Nguyen & Derek Blankenship & Clarence B Creec, 2012. "Host Immune Transcriptional Profiles Reflect the Variability in Clinical Disease Manifestations in Patients with Staphylococcus aureus Infections," PLOS ONE, Public Library of Science, vol. 7(4), pages 1-11, April.
    11. Qi Su & Qin Liu & Raphaela Iris Lau & Jingwan Zhang & Zhilu Xu & Yun Kit Yeoh & Thomas W. H. Leung & Whitney Tang & Lin Zhang & Jessie Q. Y. Liang & Yuk Kam Yau & Jiaying Zheng & Chengyu Liu & Mengjin, 2022. "Faecal microbiome-based machine learning for multi-class disease diagnosis," Nature Communications, Nature, vol. 13(1), pages 1-8, December.
    12. Nazifa Ahmed Moumi & Badhan Das & Zarin Tasnim Promi & Nishat Anjum Bristy & Md Shamsuzzoha Bayzid, 2019. "Quartet-based inference of cell differentiation trees from ChIP-Seq histone modification data," PLOS ONE, Public Library of Science, vol. 14(9), pages 1-25, September.
    13. Raihan K Uddin & Shiva M Singh, 2013. "Hippocampal Gene Expression Meta-Analysis Identifies Aging and Age-Associated Spatial Learning Impairment (ASLI) Genes and Pathways," PLOS ONE, Public Library of Science, vol. 8(7), pages 1-16, July.
    14. Kejian Wang & Jiazhi Sun & Shufeng Zhou & Chunling Wan & Shengying Qin & Can Li & Lin He & Lun Yang, 2013. "Prediction of Drug-Target Interactions for Drug Repositioning Only Based on Genomic Expression Similarity," PLOS Computational Biology, Public Library of Science, vol. 9(11), pages 1-9, November.

    More about this item

    Statistics

    Access and download statistics

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:plo:pone00:0301912. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: plosone (email available below). General contact details of provider: https://journals.plos.org/plosone/ .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.