PatternMarkers & GWCoGAPS for novel data-driven biomarkers via whole transcriptome NMF
Genevieve L Stein-O’Brien,
Jacob L Carey,
Alexander V Favorov,
Daria A Gaykalova,
Ronald D McKay,
Michael F. Ochs,
Elana J. Fertig
Posted 26 Oct 2016
bioRxiv DOI: 10.1101/083717 (published DOI: 10.1093/bioinformatics/btx058)
Posted 26 Oct 2016
Non-negative Matrix Factorization (NMF) algorithms associate gene expression with biological processes (e.g., time-course dynamics or disease subtypes). Compared with univariate associations, the relative weights of NMF solutions can obscure biomarkers. Therefore, we developed a novel PatternMarkers statistic to extract genes for biological validation and enhanced visualization of NMF results. Finding novel and unbiased gene markers with PatternMarkers requires whole-genome data. However, NMF algorithms typically do not converge for the tens of thousands of genes in genome-wide profiling. Therefore, we also developed Genome-Wide CoGAPS Analysis in Parallel Sets (GWCoGAPS), the first robust whole genome Bayesian NMF using the sparse, MCMC algorithm, CoGAPS. This software contains analytic and visualization tools including a Shiny web application, patternMatcher, which are generalized for any NMF. Using these tools, we find granular brain-region and cell-type specific signatures with corresponding biomarkers in GTex data, illustrating GWCoGAPS and patternMarkers ascertainment of data-driven biomarkers from whole-genome data. Availability: PatternMarkers & GWCoGAPS are in the CoGAPS Bioconductor package (3.5) under the GPL license.
- Downloaded 635 times
- Download rankings, all-time:
- Site-wide: 27,643 out of 103,919
- In bioinformatics: 3,711 out of 9,474
- Year to date:
- Site-wide: 63,643 out of 103,919
- Since beginning of last month:
- Site-wide: 84,173 out of 103,919
Downloads over time
Distribution of downloads per paper, site-wide
- 18 Dec 2019: We're pleased to announce PanLingua, a new tool that enables you to search for machine-translated bioRxiv preprints using more than 100 different languages.
- 21 May 2019: PLOS Biology has published a community page about Rxivist.org and its design.
- 10 May 2019: The paper analyzing the Rxivist dataset has been published at eLife.
- 1 Mar 2019: We now have summary statistics about bioRxiv downloads and submissions.
- 8 Feb 2019: Data from Altmetric is now available on the Rxivist details page for every preprint. Look for the "donut" under the download metrics.
- 30 Jan 2019: preLights has featured the Rxivist preprint and written about our findings.
- 22 Jan 2019: Nature just published an article about Rxivist and our data.
- 13 Jan 2019: The Rxivist preprint is live!