ctheodoris/Genecorpus-30M download history

ctheodoris/Genecorpus-30M is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 1,054 times (198 in the last 7 days), and 28,864 times in total. It ranks #18,593 among datasets by monthly downloads.

Dataset Card for Genecorpus-30M Dataset Description Point of Contact: christina.theodoris@gladstone.ucsf.edu Dataset Summary We assembled a large-scale pretraining corpus, Genecorpus-30M, comprised of ~30 million human single cell transcriptomes from a broad range

Models trained on Genecorpus-30M

16 models list it as training data.

Open ctheodoris/Genecorpus-30M on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.