cturan/turkish-synthetic-corpus download history

cturan/turkish-synthetic-corpus is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 46 times (7 in the last 7 days), and 877 times in total. It ranks #197,310 among datasets by monthly downloads.

Turkish Synthetic Corpus A synthetic Turkish text corpus with 1,871,131 documents, designed for Turkish language model training. About Inspired by HuggingFaceTB/smollm-corpus. Questions and prompts were sourced from the SmolLM Corpus pipeline; a language model then generated lo

Open cturan/turkish-synthetic-corpus on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.