IDEAS home Printed from https://ideas.repec.org/a/nat/natcom/v15y2024i1d10.1038_s41467-024-44781-7.html
   My bibliography  Save this article

Human whole-exome genotype data for Alzheimer’s disease

Author

Listed:
  • Yuk Yee Leung

    (University of Pennsylvania)

  • Adam C. Naj

    (University of Pennsylvania
    University of Pennsylvania)

  • Yi-Fan Chou

    (University of Pennsylvania)

  • Otto Valladares

    (University of Pennsylvania)

  • Michael Schmidt

    (Miller School of Medicine, University of Miami
    University of Miami)

  • Kara Hamilton-Nelson

    (Miller School of Medicine, University of Miami
    University of Miami)

  • Nicholas Wheeler

    (Case Western Reserve University
    Case Western Reserve University)

  • Honghuang Lin

    (UMass Chan Medical School)

  • Prabhakaran Gangadharan

    (University of Pennsylvania)

  • Liming Qu

    (University of Pennsylvania)

  • Kaylyn Clark

    (University of Pennsylvania)

  • Amanda B. Kuzma

    (University of Pennsylvania)

  • Wan-Ping Lee

    (University of Pennsylvania)

  • Laura Cantwell

    (University of Pennsylvania)

  • Heather Nicaretta

    (University of Pennsylvania)

  • Jonathan Haines

    (Case Western Reserve University
    Case Western Reserve University)

  • Lindsay Farrer

    (Boston University Chobanian & Avedisian School of Medicine
    Boston University School of Public Health)

  • Sudha Seshadri

    (Boston University School of Medicine
    University of Texas Health Sciences Center)

  • Zoran Brkanac

    (University of Washington)

  • Carlos Cruchaga

    (Washington University School of Medicine)

  • Margaret Pericak-Vance

    (Miller School of Medicine, University of Miami
    University of Miami)

  • Richard P. Mayeux

    (Columbia University and the New York Presbyterian Hospital)

  • William S. Bush

    (Case Western Reserve University
    Case Western Reserve University)

  • Anita Destefano

    (Boston University School of Public Health
    Boston University School of Medicine)

  • Eden Martin

    (Miller School of Medicine, University of Miami
    University of Miami)

  • Gerard D. Schellenberg

    (University of Pennsylvania)

  • Li-San Wang

    (University of Pennsylvania)

Abstract

The heterogeneity of the whole-exome sequencing (WES) data generation methods present a challenge to a joint analysis. Here we present a bioinformatics strategy for joint-calling 20,504 WES samples collected across nine studies and sequenced using ten capture kits in fourteen sequencing centers in the Alzheimer’s Disease Sequencing Project. The joint-genotype called variant-called format (VCF) file contains only positions within the union of capture kits. The VCF was then processed specifically to account for the batch effects arising from the use of different capture kits from different studies. We identified 8.2 million autosomal variants. 96.82% of the variants are high-quality, and are located in 28,579 Ensembl transcripts. 41% of the variants are intronic and 1.8% of the variants are with CADD > 30, indicating they are of high predicted pathogenicity. Here we show our new strategy can generate high-quality data from processing these diversely generated WES samples. The improved ability to combine data sequenced in different batches benefits the whole genomics research community.

Suggested Citation

  • Yuk Yee Leung & Adam C. Naj & Yi-Fan Chou & Otto Valladares & Michael Schmidt & Kara Hamilton-Nelson & Nicholas Wheeler & Honghuang Lin & Prabhakaran Gangadharan & Liming Qu & Kaylyn Clark & Amanda B., 2024. "Human whole-exome genotype data for Alzheimer’s disease," Nature Communications, Nature, vol. 15(1), pages 1-15, December.
  • Handle: RePEc:nat:natcom:v:15:y:2024:i:1:d:10.1038_s41467-024-44781-7
    DOI: 10.1038/s41467-024-44781-7
    as

    Download full text from publisher

    File URL: https://www.nature.com/articles/s41467-024-44781-7
    File Function: Abstract
    Download Restriction: no

    File URL: https://libkey.io/10.1038/s41467-024-44781-7?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    References listed on IDEAS

    as
    1. Nick Patterson & Alkes L Price & David Reich, 2006. "Population Structure and Eigenanalysis," PLOS Genetics, Public Library of Science, vol. 2(12), pages 1-20, December.
    2. Allison A. Regier & Yossi Farjoun & David E. Larson & Olga Krasheninina & Hyun Min Kang & Daniel P. Howrigan & Bo-Juen Chen & Manisha Kher & Eric Banks & Darren C. Ames & Adam C. English & Heng Li & J, 2018. "Functional equivalence of genome sequencing analysis pipelines enables harmonized variant calling across human genetics projects," Nature Communications, Nature, vol. 9(1), pages 1-8, December.
    3. Monkol Lek & Konrad J. Karczewski & Eric V. Minikel & Kaitlin E. Samocha & Eric Banks & Timothy Fennell & Anne H. O’Donnell-Luria & James S. Ware & Andrew J. Hill & Beryl B. Cummings & Taru Tukiainen , 2016. "Analysis of protein-coding genetic variation in 60,706 humans," Nature, Nature, vol. 536(7616), pages 285-291, August.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Gyaneshwer Chaubey & Anurag Kadian & Saroj Bala & Vadlamudi Raghavendra Rao, 2015. "Genetic Affinity of the Bhil, Kol and Gond Mentioned in Epic Ramayana," PLOS ONE, Public Library of Science, vol. 10(6), pages 1-11, June.
    2. Estavoyer, Maxime & François, Olivier, 2022. "Theoretical analysis of principal components in an umbrella model of intraspecific evolution," Theoretical Population Biology, Elsevier, vol. 148(C), pages 11-21.
    3. Hyosik Jang & Ian M Ehrenreich, 2012. "Genome-Wide Characterization of Genetic Variation in the Unicellular, Green Alga Chlamydomonas reinhardtii," PLOS ONE, Public Library of Science, vol. 7(7), pages 1-9, July.
    4. Xiaofeng Cai & Xuepeng Sun & Chenxi Xu & Honghe Sun & Xiaoli Wang & Chenhui Ge & Zhonghua Zhang & Quanxi Wang & Zhangjun Fei & Chen Jiao & Quanhua Wang, 2021. "Genomic analyses provide insights into spinach domestication and the genetic basis of agronomic traits," Nature Communications, Nature, vol. 12(1), pages 1-12, December.
    5. Lee, Anthony J. & Hibbs, Courtney & Wright, Margaret J. & Martin, Nicholas G. & Keller, Matthew C. & Zietsch, Brendan P., 2017. "Assessing the accuracy of perceptions of intelligence based on heritable facial features," Intelligence, Elsevier, vol. 64(C), pages 1-8.
    6. Thompson Katherine L. & Linnen Catherine R. & Kubatko Laura, 2016. "Tree-based quantitative trait mapping in the presence of external covariates," Statistical Applications in Genetics and Molecular Biology, De Gruyter, vol. 15(6), pages 473-490, December.
    7. Ruoyu Tian & Tian Ge & Hyeokmoon Kweon & Daniel B. Rocha & Max Lam & Jimmy Z. Liu & Kritika Singh & Daniel F. Levey & Joel Gelernter & Murray B. Stein & Ellen A. Tsai & Hailiang Huang & Christopher F., 2024. "Whole-exome sequencing in UK Biobank reveals rare genetic architecture for depression," Nature Communications, Nature, vol. 15(1), pages 1-12, December.
    8. Birgit Burkhardt & Ulf Michgehl & Jonas Rohde & Tabea Erdmann & Philipp Berning & Katrin Reutter & Marius Rohde & Arndt Borkhardt & Thomas Burmeister & Sandeep Dave & Alexandar Tzankov & Martin Dugas , 2022. "Clinical relevance of molecular characteristics in Burkitt lymphoma differs according to age," Nature Communications, Nature, vol. 13(1), pages 1-12, December.
    9. Jacobo Pardo-Seco & Alberto Gómez-Carballa & Jorge Amigo & Federico Martinón-Torres & Antonio Salas, 2014. "A Genome-Wide Study of Modern-Day Tuscans: Revisiting Herodotus's Theory on the Origin of the Etruscans," PLOS ONE, Public Library of Science, vol. 9(9), pages 1-11, September.
    10. Ilja M Nolte & Chris Wallace & Stephen J Newhouse & Daryl Waggott & Jingyuan Fu & Nicole Soranzo & Rhian Gwilliam & Panos Deloukas & Irina Savelieva & Dongling Zheng & Chrysoula Dalageorgou & Martin F, 2009. "Common Genetic Variation Near the Phospholamban Gene Is Associated with Cardiac Repolarisation: Meta-Analysis of Three Genome-Wide Association Studies," PLOS ONE, Public Library of Science, vol. 4(7), pages 1-10, July.
    11. Hoicheong Siu & Li Jin & Momiao Xiong, 2012. "Manifold Learning for Human Population Structure Studies," PLOS ONE, Public Library of Science, vol. 7(1), pages 1-18, January.
    12. Elodie Persyn & Richard Redon & Lise Bellanger & Christian Dina, 2018. "The impact of a fine-scale population stratification on rare variant association test results," PLOS ONE, Public Library of Science, vol. 13(12), pages 1-17, December.
    13. Andre Krumel Portella & Afroditi Papantoni & Catherine Paquet & Spencer Moore & Keri Shiels Rosch & Stewart Mostofsky & Richard S Lee & Kimberly R Smith & Robert Levitan & Patricia Pelufo Silveira & S, 2020. "Predicted DRD4 prefrontal gene expression moderates snack intake and stress perception in response to the environment in adolescents," PLOS ONE, Public Library of Science, vol. 15(6), pages 1-20, June.
    14. Maria Stahl Madsen & Marjoleine F. Broekema & Martin Rønn Madsen & Arjen Koppen & Anouska Borgman & Cathrin Gräwe & Elisabeth G. K. Thomsen & Denise Westland & Mariette E. G. Kranendonk & Marian Groot, 2022. "PPARγ lipodystrophy mutants reveal intermolecular interactions required for enhancer activation," Nature Communications, Nature, vol. 13(1), pages 1-19, December.
    15. Lindsay Fernández-Rhodes & Jennifer R Malinowski & Yujie Wang & Ran Tao & Nathan Pankratz & Janina M Jeff & Sachiko Yoneyama & Cara L Carty & V Wendy Setiawan & Loic Le Marchand & Christopher Haiman &, 2018. "The genetic underpinnings of variation in ages at menarche and natural menopause among women from the multi-ethnic Population Architecture using Genomics and Epidemiology (PAGE) Study: A trans-ethnic ," PLOS ONE, Public Library of Science, vol. 13(7), pages 1-21, July.
    16. Mark W. Youngblood & Zeynep Erson-Omay & Chang Li & Hinda Najem & Süleyman Coșkun & Evgeniya Tyrtova & Julio D. Montejo & Danielle F. Miyagishima & Tanyeri Barak & Sayoko Nishimura & Akdes Serin Harma, 2023. "Super-enhancer hijacking drives ectopic expression of hedgehog pathway ligands in meningiomas," Nature Communications, Nature, vol. 14(1), pages 1-16, December.
    17. Peña-Malavera Andrea & Bruno Cecilia & Balzarini Monica & Fernandez Elmer, 2014. "Comparison of algorithms to infer genetic population structure from unlinked molecular markers," Statistical Applications in Genetics and Molecular Biology, De Gruyter, vol. 13(4), pages 1-12, August.
    18. Chi-Chun Liu & David Witonsky & Anna Gosling & Ju Hyeon Lee & Harald Ringbauer & Richard Hagan & Nisha Patel & Raphaela Stahl & John Novembre & Mark Aldenderfer & Christina Warinner & Anna Di Rienzo &, 2022. "Ancient genomes from the Himalayas illuminate the genetic history of Tibetans and their Tibeto-Burman speaking neighbors," Nature Communications, Nature, vol. 13(1), pages 1-14, December.
    19. James T. Topham & Erica S. Tsang & Joanna M. Karasinska & Andrew Metcalfe & Hassan Ali & Steve E. Kalloger & Veronika Csizmok & Laura M. Williamson & Emma Titmuss & Karina Nielsen & Gian Luca Negri & , 2022. "Integrative analysis of KRAS wildtype metastatic pancreatic ductal adenocarcinoma reveals mutation and expression-based similarities to cholangiocarcinoma," Nature Communications, Nature, vol. 13(1), pages 1-13, December.
    20. Gad Abraham & Michael Inouye, 2014. "Fast Principal Component Analysis of Large-Scale Genome-Wide Data," PLOS ONE, Public Library of Science, vol. 9(4), pages 1-5, April.

    More about this item

    Statistics

    Access and download statistics

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:nat:natcom:v:15:y:2024:i:1:d:10.1038_s41467-024-44781-7. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Sonal Shukla or Springer Nature Abstracting and Indexing (email available below). General contact details of provider: http://www.nature.com .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.