linagora/fineweb2_Tunisian_Arabic download history

linagora/fineweb2_Tunisian_Arabic is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 60 times (10 in the last 7 days), and 488 times in total. It ranks #162,992 among datasets by monthly downloads.

This is the Tunisian Arabic Portion of The FineWeb2 Dataset. This dataset contains a rich collection of text in Tunisian Arabic (ISO 639-3: aeb), a widely spoken dialect within the Afro-Asiatic language family. It serves as a valuable resource for NLP development and linguistic research focused on

Models trained on fineweb2_Tunisian_Arabic

1 models list it as training data.

Open linagora/fineweb2_Tunisian_Arabic on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.