Henrychur/MMedC download history

Henrychur/MMedC is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 449 times (41 in the last 7 days), and 4,526 times in total. It ranks #36,061 among datasets by monthly downloads.

MMedC 💻Github Repo 🖨️arXiv Paper The official pre-training dataset for "Towards Building Multilingual Language Model for Medicine". News We add Arabic and German corpus to MMedC. Introduction This repo contains MMedC, a multilingual medical corpus with 25.5 billi

Models trained on MMedC

16 models list it as training data.

Open Henrychur/MMedC on Hugging Face