NAMAA-Space/QariOCR-v0.3-markdown-mixed-dataset download history
NAMAA-Space/QariOCR-v0.3-markdown-mixed-dataset is a text to image dataset on the Hugging Face Hub. In the last 30 days it was downloaded 530 times (27 in the last 7 days), and 3,364 times in total. It ranks #31,690 among datasets by monthly downloads.
QARI Markdown Mixed Dataset 📋 Dataset Summary The QARI v0.3 Markdown Mixed Dataset is a specialized synthetic dataset designed for training Arabic OCR models with a focus on complex document layouts and HTML structure understanding. This dataset is part of
Models trained on QariOCR-v0.3-markdown-mixed-dataset
4 models list it as training data.
- NAMAA-Space/Qari-OCR-v0.3-VL-2B-Instruct 2.8K downloads in 30 days
- mradermacher/Qari-OCR-0.3-SNAPSHOT-VL-2B-Instruct-merged-GGUF 1.1K downloads in 30 days
- FatimahEmadEldin/Waraqon-v3-Arabic-OCR-HTML-Qari 25 downloads in 30 days
- mo1998/arabic-ocr-qwen2.5-vl 0 downloads in 30 days
Open NAMAA-Space/QariOCR-v0.3-markdown-mixed-dataset on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.