linagora/fineweb2_Tunisian_Arabic download history
linagora/fineweb2_Tunisian_Arabic is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 60 times (10 in the last 7 days), and 488 times in total. It ranks #162,992 among datasets by monthly downloads.
This is the Tunisian Arabic Portion of The FineWeb2 Dataset. This dataset contains a rich collection of text in Tunisian Arabic (ISO 639-3: aeb), a widely spoken dialect within the Afro-Asiatic language family. It serves as a valuable resource for NLP development and linguistic research focused on
Models trained on fineweb2_Tunisian_Arabic
1 models list it as training data.
- messalti/MagharibiBERT-v1 456 downloads in 30 days
Open linagora/fineweb2_Tunisian_Arabic on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.