Corpus-NZ/English download history

Corpus-NZ/English is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 13 times (5 in the last 7 days), and 46 times in total. It ranks #491,332 among datasets by monthly downloads.

Synthetic English Language Acquisition Dataset (3GB) A structured, 3GB synthetic CSV dataset generated to assist in pretraining or fine-tuning Language Models (LLMs) on core English syntax, vocabulary, narrative structures, and explicit grammar rules. Dataset Structure

Open Corpus-NZ/English on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.