loay/arabic-ocr-synthetic-scans-faker-300k download history
loay/arabic-ocr-synthetic-scans-faker-300k is an image to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 387 times (46 in the last 7 days), and 4,982 times in total. It ranks #40,356 among datasets by monthly downloads.
Arabic OCR Synthetic Scans (Faker 300k) A large-scale synthetic dataset of ~300,000 Arabic book pages generated to mimic real-world scanning imperfections. Designed for training Vision Language Models (VLMs) and OCR engines on structural layout analysis, font recognition, and document deg
Models trained on arabic-ocr-synthetic-scans-faker-300k
2 models list it as training data.
- mradermacher/Arabic-Qwen3.5-OCR-v4-GGUF 767 downloads in 30 days
- sherif1313/Arabic-Qwen3.5-OCR-v4 336 downloads in 30 days
Open loay/arabic-ocr-synthetic-scans-faker-300k on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.