tunis-ai/tunisian-msa-parallel-corpus-evaluated download history
tunis-ai/tunisian-msa-parallel-corpus-evaluated is a translation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 26 times (10 in the last 7 days), and 524 times in total. It ranks #306,850 among datasets by monthly downloads.
Dataset Description This dataset is a synthetic parallel corpus of Tunisian Arabic (aeb) and Modern Standard Arabic (arb). It was created with a rigorous multi-stage pipeline to maximize quality and reproducibility, addressing the scarcity of high-quality resources for Tunisian Arabic
Open tunis-ai/tunisian-msa-parallel-corpus-evaluated on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.