IDEAS home Printed from https://ideas.repec.org/a/plo/pgen00/1003594.html
   My bibliography  Save this article

DeepSAGE Reveals Genetic Variants Associated with Alternative Polyadenylation and Expression of Coding and Non-coding Transcripts

Author

Listed:
  • Daria V Zhernakova
  • Eleonora de Klerk
  • Harm-Jan Westra
  • Anastasios Mastrokolias
  • Shoaib Amini
  • Yavuz Ariyurek
  • Rick Jansen
  • Brenda W Penninx
  • Jouke J Hottenga
  • Gonneke Willemsen
  • Eco J de Geus
  • Dorret I Boomsma
  • Jan H Veldink
  • Leonard H van den Berg
  • Cisca Wijmenga
  • Johan T den Dunnen
  • Gert-Jan B van Ommen
  • Peter A C 't Hoen
  • Lude Franke

Abstract

Many disease-associated variants affect gene expression levels (expression quantitative trait loci, eQTLs) and expression profiling using next generation sequencing (NGS) technology is a powerful way to detect these eQTLs. We analyzed 94 total blood samples from healthy volunteers with DeepSAGE to gain specific insight into how genetic variants affect the expression of genes and lengths of 3′-untranslated regions (3′-UTRs). We detected previously unknown cis-eQTL effects for GWAS hits in disease- and physiology-associated traits. Apart from cis-eQTLs that are typically easily identifiable using microarrays or RNA-sequencing, DeepSAGE also revealed many cis-eQTLs for antisense and other non-coding transcripts, often in genomic regions containing retrotransposon-derived elements. We also identified and confirmed SNPs that affect the usage of alternative polyadenylation sites, thereby potentially influencing the stability of messenger RNAs (mRNA). We then combined the power of RNA-sequencing with DeepSAGE by performing a meta-analysis of three datasets, leading to the identification of many more cis-eQTLs. Our results indicate that DeepSAGE data is useful for eQTL mapping of known and unknown transcripts, and for identifying SNPs that affect alternative polyadenylation. Because of the inherent differences between DeepSAGE and RNA-sequencing, our complementary, integrative approach leads to greater insight into the molecular consequences of many disease-associated variants.Author Summary: Many genetic variants that are associated with diseases also affect gene expression levels. We used a next generation sequencing approach targeting 3′ transcript ends (DeepSAGE) to gain specific insight into how genetic variants affect the expression of genes and the usage and length of 3′-untranslated regions. We detected many associations for antisense and other non-coding transcripts, often in genomic regions containing retrotransposon-derived elements. Some of these variants are also associated with disease. We also identified and confirmed variants that affect the usage of alternative polyadenylation sites, thereby potentially influencing the stability of mRNAs. We conclude that DeepSAGE is useful for detecting eQTL effects on both known and unknown transcripts, and for identifying variants that affect alternative polyadenylation.

Suggested Citation

  • Daria V Zhernakova & Eleonora de Klerk & Harm-Jan Westra & Anastasios Mastrokolias & Shoaib Amini & Yavuz Ariyurek & Rick Jansen & Brenda W Penninx & Jouke J Hottenga & Gonneke Willemsen & Eco J de Ge, 2013. "DeepSAGE Reveals Genetic Variants Associated with Alternative Polyadenylation and Expression of Coding and Non-coding Transcripts," PLOS Genetics, Public Library of Science, vol. 9(6), pages 1-15, June.
  • Handle: RePEc:plo:pgen00:1003594
    DOI: 10.1371/journal.pgen.1003594
    as

    Download full text from publisher

    File URL: https://journals.plos.org/plosgenetics/article?id=10.1371/journal.pgen.1003594
    Download Restriction: no

    File URL: https://journals.plos.org/plosgenetics/article/file?id=10.1371/journal.pgen.1003594&type=printable
    Download Restriction: no

    File URL: https://libkey.io/10.1371/journal.pgen.1003594?utm_source=ideas
    LibKey link: if access is restricted and if your library uses this service, LibKey will redirect you to where you can use your library subscription to access this item
    ---><---

    References listed on IDEAS

    as
    1. Vivian G. Cheung & Richard S. Spielman & Kathryn G. Ewens & Teresa M. Weber & Michael Morley & Joshua T. Burdick, 2005. "Mapping determinants of human gene expression by regional and genome-wide association," Nature, Nature, vol. 437(7063), pages 1365-1369, October.
    2. Joseph K. Pickrell & John C. Marioni & Athma A. Pai & Jacob F. Degner & Barbara E. Engelhardt & Everlyne Nkadori & Jean-Baptiste Veyrieras & Matthew Stephens & Yoav Gilad & Jonathan K. Pritchard, 2010. "Understanding mechanisms underlying human gene expression variation with RNA sequencing," Nature, Nature, vol. 464(7289), pages 768-772, April.
    3. Stephen B. Montgomery & Micha Sammeth & Maria Gutierrez-Arcelus & Radoslaw P. Lach & Catherine Ingle & James Nisbett & Roderic Guigo & Emmanouil T. Dermitzakis, 2010. "Transcriptome genetics using second generation sequencing in a Caucasian population," Nature, Nature, vol. 464(7289), pages 773-777, April.
    Full references (including those not matched with items on IDEAS)

    Citations

    Citations are extracted by the CitEc Project, subscribe to its RSS feed for this item.
    as


    Cited by:

    1. Urmo Võsa & Tõnu Esko & Silva Kasela & Tarmo Annilo, 2015. "Altered Gene Expression Associated with microRNA Binding Site Polymorphisms," PLOS ONE, Public Library of Science, vol. 10(10), pages 1-24, October.

    Most related items

    These are the items that most often cite the same works as this one and are cited by the same works as this one.
    1. Alexandra C Nica & Leopold Parts & Daniel Glass & James Nisbet & Amy Barrett & Magdalena Sekowska & Mary Travers & Simon Potter & Elin Grundberg & Kerrin Small & Åsa K Hedman & Veronique Bataille & Jo, 2011. "The Architecture of Gene Regulatory Variation across Multiple Human Tissues: The MuTHER Study," PLOS Genetics, Public Library of Science, vol. 7(2), pages 1-9, February.
    2. Barbara E Stranger & Stephen B Montgomery & Antigone S Dimas & Leopold Parts & Oliver Stegle & Catherine E Ingle & Magda Sekowska & George Davey Smith & David Evans & Maria Gutierrez-Arcelus & Alkes P, 2012. "Patterns of Cis Regulatory Variation in Diverse Human Populations," PLOS Genetics, Public Library of Science, vol. 8(4), pages 1-13, April.
    3. Jin Hyun Ju & Sushila A Shenoy & Ronald G Crystal & Jason G Mezey, 2017. "An independent component analysis confounding factor correction framework for identifying broad impact expression quantitative trait loci," PLOS Computational Biology, Public Library of Science, vol. 13(5), pages 1-26, May.
    4. Faisal Shahla & Tutz Gerhard, 2017. "Missing value imputation for gene expression data by tailored nearest neighbors," Statistical Applications in Genetics and Molecular Biology, De Gruyter, vol. 16(2), pages 95-106, April.
    5. Thanh Nguyen & Asim Bhatti & Samuel Yang & Saeid Nahavandi, 2016. "RNA-Seq Count Data Modelling by Grey Relational Analysis and Nonparametric Gaussian Process," PLOS ONE, Public Library of Science, vol. 11(10), pages 1-18, October.
    6. Kensuke Yamaguchi & Kazuyoshi Ishigaki & Akari Suzuki & Yumi Tsuchida & Haruka Tsuchiya & Shuji Sumitomo & Yasuo Nagafuchi & Fuyuki Miya & Tatsuhiko Tsunoda & Hirofumi Shoda & Keishi Fujio & Kazuhiko , 2022. "Splicing QTL analysis focusing on coding sequences reveals mechanisms for disease susceptibility loci," Nature Communications, Nature, vol. 13(1), pages 1-13, December.
    7. Jean Francois Lefebvre & Emilio Vello & Bing Ge & Stephen B Montgomery & Emmanouil T Dermitzakis & Tomi Pastinen & Damian Labuda, 2012. "Genotype-Based Test in Mapping Cis-Regulatory Variants from Allele-Specific Expression Data," PLOS ONE, Public Library of Science, vol. 7(6), pages 1-15, June.
    8. Sora Yoon & Seon-Young Kim & Dougu Nam, 2016. "Improving Gene-Set Enrichment Analysis of RNA-Seq Data with Small Replicates," PLOS ONE, Public Library of Science, vol. 11(11), pages 1-16, November.
    9. Yixin Fang & Yang Feng & Ming Yuan, 2014. "Regularized principal components of heritability," Computational Statistics, Springer, vol. 29(3), pages 455-465, June.
    10. Pingting Ying & Can Chen & Zequn Lu & Shuoni Chen & Ming Zhang & Yimin Cai & Fuwei Zhang & Jinyu Huang & Linyun Fan & Caibo Ning & Yanmin Li & Wenzhuo Wang & Hui Geng & Yizhuo Liu & Wen Tian & Zhiyong, 2023. "Genome-wide enhancer-gene regulatory maps link causal variants to target genes underlying human cancer risk," Nature Communications, Nature, vol. 14(1), pages 1-20, December.
    11. Kyung-Won Hong & Seok Won Jeong & Myungguen Chung & Seong Beom Cho, 2014. "Association between Expression Quantitative Trait Loci and Metabolic Traits in Two Korean Populations," PLOS ONE, Public Library of Science, vol. 9(12), pages 1-13, December.
    12. Xiaodong Cai & Juan Andrés Bazerque & Georgios B Giannakis, 2013. "Inference of Gene Regulatory Networks with Sparse Structural Equation Models Exploiting Genetic Perturbations," PLOS Computational Biology, Public Library of Science, vol. 9(5), pages 1-13, May.
    13. Ryan Abo & Gregory D Jenkins & Liewei Wang & Brooke L Fridley, 2012. "Identifying the Genetic Variation of Gene Expression Using Gene Sets: Application of Novel Gene Set eQTL Approach to PharmGKB and KEGG," PLOS ONE, Public Library of Science, vol. 7(8), pages 1-11, August.
    14. Nicoló Fusi & Oliver Stegle & Neil D Lawrence, 2012. "Joint Modelling of Confounding Factors and Prominent Genetic Regulators Provides Increased Accuracy in Genetical Genomics Studies," PLOS Computational Biology, Public Library of Science, vol. 8(1), pages 1-9, January.
    15. Bin Wang, 2020. "A Zipf-plot based normalization method for high-throughput RNA-seq data," PLOS ONE, Public Library of Science, vol. 15(4), pages 1-15, April.
    16. Jungsoo Gim & Sungho Won & Taesung Park, 2016. "LPEseq: Local-Pooled-Error Test for RNA Sequencing Experiments with a Small Number of Replicates," PLOS ONE, Public Library of Science, vol. 11(8), pages 1-15, August.
    17. Ning Jiang & Minghui Wang & Tianye Jia & Lin Wang & Lindsey Leach & Christine Hackett & David Marshall & Zewei Luo, 2011. "A Robust Statistical Method for Association-Based eQTL Analysis," PLOS ONE, Public Library of Science, vol. 6(8), pages 1-11, August.
    18. Paul C Boutros & Ivy D Moffat & Allan B Okey & Raimo Pohjanvirta, 2011. "mRNA Levels in Control Rat Liver Display Strain-Specific, Hereditary, and AHR-Dependent Components," PLOS ONE, Public Library of Science, vol. 6(7), pages 1-15, July.
    19. Eric O Johnson & Dana B Hancock & Nathan C Gaddis & Joshua L Levy & Grier Page & Scott P Novak & Cristie Glasheen & Nancy L Saccone & John P Rice & Michael P Moreau & Kimberly F Doheny & Jane M Romm &, 2015. "Novel Genetic Locus Implicated for HIV-1 Acquisition with Putative Regulatory Links to HIV Replication and Infectivity: A Genome-Wide Association Study," PLOS ONE, Public Library of Science, vol. 10(3), pages 1-15, March.
    20. Tang Clara S. & Ferreira Manuel A. R., 2012. "GENOVA: Gene Overlap Analysis of GWAS Results," Statistical Applications in Genetics and Molecular Biology, De Gruyter, vol. 11(3), pages 1-15, February.

    More about this item

    Statistics

    Access and download statistics

    Corrections

    All material on this site has been provided by the respective publishers and authors. You can help correct errors and omissions. When requesting a correction, please mention this item's handle: RePEc:plo:pgen00:1003594. See general information about how to correct material in RePEc.

    If you have authored this item and are not yet registered with RePEc, we encourage you to do it here. This allows to link your profile to this item. It also allows you to accept potential citations to this item that we are uncertain about.

    If CitEc recognized a bibliographic reference but did not link an item in RePEc to it, you can help with this form .

    If you know of missing items citing this one, you can help us creating those links by adding the relevant references in the same way as above, for each refering item. If you are a registered author of this item, you may also want to check the "citations" tab in your RePEc Author Service profile, as there may be some citations waiting for confirmation.

    For technical questions regarding this item, or to correct its authors, title, abstract, bibliographic or download information, contact: plosgenetics (email available below). General contact details of provider: https://journals.plos.org/plosgenetics/ .

    Please note that corrections may take a couple of weeks to filter through the various RePEc services.

    IDEAS is a RePEc service. RePEc uses bibliographic data supplied by the respective publishers.