IDEAS home Printed from https://ideas.repec.org/a/eee/csdana/v180y2023ics0167947322002298.html
   My bibliography  Save this article

Jackstraw inference for AJIVE data integration

Author

Listed:
  • Yang, Xi
  • Hoadley, Katherine A.
  • Hannig, Jan
  • Marron, J.S.

Abstract

In the age of big data, data integration is a critical step especially in the understanding of how diverse data types work together and work separately. Among data integration methods, the Angle-Based Joint and Individual Variation Explained (AJIVE) approach is particularly attractive because it not only studies joint behavior but also individual behavior. Typically AJIVE scores indicate important relationships between data objects, such as clusters. An important challenge is understanding which features, i.e. variables, are associated with those relationships. This challenge is addressed by the proposal of a hypothesis test for assessing statistical significance of features. The new test is inspired by the related jackstraw method developed for Principal Component Analysis. We use a high-dimensional multi-genomic cancer data set as our strong motivation and deep illustration of the methodology.

Suggested Citation

  • Yang, Xi & Hoadley, Katherine A. & Hannig, Jan & Marron, J.S., 2023. "Jackstraw inference for AJIVE data integration," Computational Statistics & Data Analysis, Elsevier, vol. 180(C).
  • Handle: RePEc:eee:csdana:v:180:y:2023:i:c:s0167947322002298
    DOI: 10.1016/j.csda.2022.107649
    as

    Download full text from publisher

    File URL: http://www.sciencedirect.com/science/article/pii/S0167947322002298
    Download Restriction: Full text for ScienceDirect subscribers only.

    File URL: https://libkey.io/10.1016/j.csda.2022.107649?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    As the access to this document is restricted, you may want to search for a different version of it.

    References listed on IDEAS

    as
    1. Charles M. Perou & Therese Sørlie & Michael B. Eisen & Matt van de Rijn & Stefanie S. Jeffrey & Christian A. Rees & Jonathan R. Pollack & Douglas T. Ross & Hilde Johnsen & Lars A. Akslen & Øystein Flu, 2000. "Molecular portraits of human breast tumours," Nature, Nature, vol. 406(6797), pages 747-752, August.
    2. Feng, Qing & Jiang, Meilei & Hannig, Jan & Marron, J.S., 2018. "Angle-based joint and individual variation explained," Journal of Multivariate Analysis, Elsevier, vol. 166(C), pages 241-265.
    Full references (including those not matched with items on IDEAS)

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. María Elena Martínez & Jonathan T Unkart & Li Tao & Candyce H Kroenke & Richard Schwab & Ian Komenaka & Scarlett Lin Gomez, 2017. "Prognostic significance of marital status in breast cancer survival: A population-based study," PLOS ONE, Public Library of Science, vol. 12(5), pages 1-14, May.
    2. Yishai Shimoni, 2018. "Association between expression of random gene sets and survival is evident in multiple cancer types and may be explained by sub-classification," PLOS Computational Biology, Public Library of Science, vol. 14(2), pages 1-15, February.
    3. Yoo-Ah Kim & Stefan Wuchty & Teresa M Przytycka, 2011. "Identifying Causal Genes and Dysregulated Pathways in Complex Diseases," PLOS Computational Biology, Public Library of Science, vol. 7(3), pages 1-13, March.
    4. Radhakrishnan Nagarajan & Marco Scutari, 2013. "Impact of Noise on Molecular Network Inference," PLOS ONE, Public Library of Science, vol. 8(12), pages 1-12, December.
    5. R Joseph Bender & Feilim Mac Gabhann, 2013. "Expression of VEGF and Semaphorin Genes Define Subgroups of Triple Negative Breast Cancer," PLOS ONE, Public Library of Science, vol. 8(5), pages 1-15, May.
    6. Deepak Poduval & Zuzana Sichmanova & Anne Hege Straume & Per Eystein Lønning & Stian Knappskog, 2020. "The novel microRNAs hsa-miR-nov7 and hsa-miR-nov3 are over-expressed in locally advanced breast cancer," PLOS ONE, Public Library of Science, vol. 15(4), pages 1-23, April.
    7. Zhiguang Huo & Li Zhu & Tianzhou Ma & Hongcheng Liu & Song Han & Daiqing Liao & Jinying Zhao & George Tseng, 2020. "Two-Way Horizontal and Vertical Omics Integration for Disease Subtype Discovery," Statistics in Biosciences, Springer;International Chinese Statistical Association, vol. 12(1), pages 1-22, April.
    8. Markus Ringnér & Erik Fredlund & Jari Häkkinen & Åke Borg & Johan Staaf, 2011. "GOBO: Gene Expression-Based Outcome for Breast Cancer Online," PLOS ONE, Public Library of Science, vol. 6(3), pages 1-11, March.
    9. Casey S Greene & Olga G Troyanskaya, 2012. "Chapter 2: Data-Driven View of Disease Biology," PLOS Computational Biology, Public Library of Science, vol. 8(12), pages 1-8, December.
    10. Mark Reimers, 2010. "Making Informed Choices about Microarray Data Analysis," PLOS Computational Biology, Public Library of Science, vol. 6(5), pages 1-7, May.
    11. Alan A. Arslan & Yian Zhang & Nedim Durmus & Sultan Pehlivan & Adrienne Addessi & Freya Schnabel & Yongzhao Shao & Joan Reibman, 2021. "Breast Cancer Characteristics in the Population of Survivors Participating in the World Trade Center Environmental Health Center Program 2002–2019," IJERPH, MDPI, vol. 18(14), pages 1-11, July.
    12. Sandra M. Rocha & Sílvia Socorro & Luís A. Passarinha & Cláudio J. Maia, 2022. "Comprehensive Landscape of STEAP Family Members Expression in Human Cancers: Unraveling the Potential Usefulness in Clinical Practice Using Integrated Bioinformatics Analysis," Data, MDPI, vol. 7(5), pages 1-48, May.
    13. J. S. Marron, 2019. "Comments on: Data science, big data and statistics," TEST: An Official Journal of the Spanish Society of Statistics and Operations Research, Springer;Sociedad de Estadística e Investigación Operativa, vol. 28(2), pages 342-344, June.
    14. Martin H van Vliet & Christiaan N Klijn & Lodewyk F A Wessels & Marcel J T Reinders, 2007. "Module-Based Outcome Prediction Using Breast Cancer Compendia," PLOS ONE, Public Library of Science, vol. 2(10), pages 1-10, October.
    15. Sung Gwe Ahn & Minkyung Lee & Tae Joo Jeon & Kyunghwa Han & Hak Min Lee & Seung Ah Lee & Young Hoon Ryu & Eun Ju Son & Joon Jeong, 2014. "[18F]-Fluorodeoxyglucose Positron Emission Tomography Can Contribute to Discriminate Patients with Poor Prognosis in Hormone Receptor-Positive Breast Cancer," PLOS ONE, Public Library of Science, vol. 9(8), pages 1-7, August.
    16. Erhan Bilal & Janusz Dutkowski & Justin Guinney & In Sock Jang & Benjamin A Logsdon & Gaurav Pandey & Benjamin A Sauerwine & Yishai Shimoni & Hans Kristian Moen Vollan & Brigham H Mecham & Oscar M Rue, 2013. "Improving Breast Cancer Survival Analysis through Competition-Based Multidimensional Modeling," PLOS Computational Biology, Public Library of Science, vol. 9(5), pages 1-16, May.
    17. Maurizio Callari & Antonio Lembo & Giampaolo Bianchini & Valeria Musella & Vera Cappelletti & Luca Gianni & Maria Grazia Daidone & Paolo Provero, 2014. "Accurate Data Processing Improves the Reliability of Affymetrix Gene Expression Profiles from FFPE Samples," PLOS ONE, Public Library of Science, vol. 9(1), pages 1-10, January.
    18. Silje Kjølle & Kenneth Finne & Even Birkeland & Vandana Ardawatia & Ingeborg Winge & Sura Aziz & Gøril Knutsvik & Elisabeth Wik & Joao A. Paulo & Heidrun Vethe & Dimitrios Kleftogiannis & Lars A. Aksl, 2023. "Hypoxia induced responses are reflected in the stromal proteome of breast cancer," Nature Communications, Nature, vol. 14(1), pages 1-16, December.
    19. Zheqi Li & Olivia McGinn & Yang Wu & Amir Bahreini & Nolan M. Priedigkeit & Kai Ding & Sayali Onkar & Caleb Lampenfeld & Carol A. Sartorius & Lori Miller & Margaret Rosenzweig & Ofir Cohen & Nikhil Wa, 2022. "ESR1 mutant breast cancers show elevated basal cytokeratins and immune activation," Nature Communications, Nature, vol. 13(1), pages 1-18, December.
    20. Leann A. Lovejoy & Craig D. Shriver & Svasti Haricharan & Rachel E. Ellsworth, 2023. "Survival Disparities in US Black Compared to White Women with Hormone Receptor Positive-HER2 Negative Breast Cancer," IJERPH, MDPI, vol. 20(4), pages 1-15, February.

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:eee:csdana:v:180:y:2023:i:c:s0167947322002298. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: Catherine Liu (email available below). General contact details of provider: http://www.elsevier.com/locate/csda .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.