We found a match
Your institution may have rights to this item. Sign in to continue.
- Title
Transfer learning for clustering single-cell RNA-seq data crossing-species and batch, case on uterine fibroids.
- Authors
Wang, Yu Mei; Sun, Yuzhi; Wang, Beiying; Wu, Zhiping; He, Xiao Ying; Zhao, Yuansong
- Abstract
Due to the high dimensionality and sparsity of the gene expression matrix in single-cell RNA-sequencing (scRNA-seq) data, coupled with significant noise generated by shallow sequencing, it poses a great challenge for cell clustering methods. While numerous computational methods have been proposed, the majority of existing approaches center on processing the target dataset itself. This approach disregards the wealth of knowledge present within other species and batches of scRNA-seq data. In light of this, our paper proposes a novel method named graph-based deep embedding clustering (GDEC) that leverages transfer learning across species and batches. GDEC integrates graph convolutional networks, effectively overcoming the challenges posed by sparse gene expression matrices. Additionally, the incorporation of DEC in GDEC enables the partitioning of cell clusters within a lower-dimensional space, thereby mitigating the adverse effects of noise on clustering outcomes. GDEC constructs a model based on existing scRNA-seq datasets and then applying transfer learning techniques to fine-tune the model using a limited amount of prior knowledge gleaned from the target dataset. This empowers GDEC to adeptly cluster scRNA-seq data cross different species and batches. Through cross-species and cross-batch clustering experiments, we conducted a comparative analysis between GDEC and conventional packages. Furthermore, we implemented GDEC on the scRNA-seq data of uterine fibroids. Compared results obtained from the Seurat package, GDEC unveiled a novel cell type (epithelial cells) and identified a notable number of new pathways among various cell types, thus underscoring the enhanced analytical capabilities of GDEC. Availability and implementation: https://github.com/YuzhiSun/GDEC/tree/main
- Subjects
UTERINE fibroids; RNA sequencing; GENE expression; EPITHELIAL cells; PRIOR learning
- Publication
Briefings in Bioinformatics, 2024, Vol 25, Issue 1, p1
- ISSN
1467-5463
- Publication type
Article
- DOI
10.1093/bib/bbad426