SLPL/naab download history

SLPL/naab is a fill mask dataset on the Hugging Face Hub. In the last 30 days it was downloaded 1,993 times (388 in the last 7 days), and 236,928 times in total. It ranks #11,326 among datasets by monthly downloads.

Huge corpora of textual data are always known to be a crucial need for training deep models such as transformer-based ones. This issue is emerging more in lower resource languages - like Farsi. We propose naab, the biggest cleaned and ready-to-use open-source textual corpus in Farsi. It contains abo

Models trained on naab

2 models list it as training data.

Open SLPL/naab on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.