scLM: automatic detection of consensus gene clusters across multiple single-cell datasets
By
Qianqian Song,
Jing Su,
Lance D. Miller,
Wei Zhang
Posted 24 Apr 2020
bioRxiv DOI: 10.1101/2020.04.22.055822
In gene expression profiling studies, including single-cell RNA-seq (scRNAseq) analyses, the identification and characterization of co-expressed genes provides critical information on cell identity and function. Gene co-expression clustering in scRNA-seq data presents certain challenges. We show that commonly used methods for single cell data are not capable of identifying co-expressed genes accurately, and produce results that substantially limit biological expectations of co-expressed genes. Herein, we present scLM, a gene co-clustering algorithm tailored to single cell data that performs well at detecting gene clusters with significant biologic context. scLM can simultaneously cluster multiple single-cell datasets, i.e. consensus clustering, enabling users to leverage single cell data from multiple sources for novel comparative analysis. scLM takes raw count data as input and preserves biological variations without being influenced by batch effects from multiple datasets. Results from both simulation data and experimental data demonstrate that scLM outperforms the existing methods with considerably improved accuracy. To illustrate the biological insights of scLM, we apply it to our in-house and public experimental scRNA-seq datasets. scLM identifies novel functional gene modules and refines cell states, which facilitates mechanism discovery and understanding of complex biosystems such as cancers. A user-friendly R package with all the key features of the scLM method is available at https://github.com/QSong-WF/scLM. ### Competing Interest Statement The authors have declared no competing interest.
Download data
- Downloaded 650 times
- Download rankings, all-time:
- Site-wide: 38,577
- In bioinformatics: 4,221
- Year to date:
- Site-wide: 21,217
- Since beginning of last month:
- Site-wide: 22,499
Altmetric data
Downloads over time
Distribution of downloads per paper, site-wide
PanLingua
News
- 27 Nov 2020: The website and API now include results pulled from medRxiv as well as bioRxiv.
- 18 Dec 2019: We're pleased to announce PanLingua, a new tool that enables you to search for machine-translated bioRxiv preprints using more than 100 different languages.
- 21 May 2019: PLOS Biology has published a community page about Rxivist.org and its design.
- 10 May 2019: The paper analyzing the Rxivist dataset has been published at eLife.
- 1 Mar 2019: We now have summary statistics about bioRxiv downloads and submissions.
- 8 Feb 2019: Data from Altmetric is now available on the Rxivist details page for every preprint. Look for the "donut" under the download metrics.
- 30 Jan 2019: preLights has featured the Rxivist preprint and written about our findings.
- 22 Jan 2019: Nature just published an article about Rxivist and our data.
- 13 Jan 2019: The Rxivist preprint is live!