Lelonthecodeur/every-language-dataset-v3 download history
Lelonthecodeur/every-language-dataset-v3 is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 404 times (183 in the last 7 days), and 404 times in total. It ranks #39,110 among datasets by monthly downloads.
Every Language Dataset V3 Next-generation synthetic multilingual dataset. Size Total: 25,000,000 Train: 24,000,000 Validation: 500,000 Test: 500,000 Diversity Human-language catalog: 171 language codes. Programming languages: 50. Task families: convers
Open Lelonthecodeur/every-language-dataset-v3 on Hugging Face