ctheodoris/Genecorpus-30M download history
ctheodoris/Genecorpus-30M is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 1,054 times (198 in the last 7 days), and 28,864 times in total. It ranks #18,593 among datasets by monthly downloads.
Dataset Card for Genecorpus-30M Dataset Description Point of Contact: christina.theodoris@gladstone.ucsf.edu Dataset Summary We assembled a large-scale pretraining corpus, Genecorpus-30M, comprised of ~30 million human single cell transcriptomes from a broad range
Models trained on Genecorpus-30M
16 models list it as training data.
- ctheodoris/Geneformer 4.6K downloads in 30 days
- nvidia/geneformer_V2_316M 320 downloads in 30 days
- nvidia/geneformer_V2_104M 299 downloads in 30 days
- nvidia/geneformer_V1_10M 284 downloads in 30 days
- nvidia/geneformer_V2_104M_CLcancer 275 downloads in 30 days
- Iambackup/geneformer_V2_316M 86 downloads in 30 days
- vpcola/Geneformer 28 downloads in 30 days
- pengxunduo/Geneformer-RNAquarium 22 downloads in 30 days
- Vinteliar/Geneformer 13 downloads in 30 days
- VirtualCell2025/Geneformer 10 downloads in 30 days
- Anler777/Geneformer 9 downloads in 30 days
- jinbo1129/try2 8 downloads in 30 days
Open ctheodoris/Genecorpus-30M on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.