loay/arabic-ocr-synthetic-scans-faker-300k download history

loay/arabic-ocr-synthetic-scans-faker-300k is an image to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 387 times (46 in the last 7 days), and 4,982 times in total. It ranks #40,356 among datasets by monthly downloads.

Arabic OCR Synthetic Scans (Faker 300k) A large-scale synthetic dataset of ~300,000 Arabic book pages generated to mimic real-world scanning imperfections. Designed for training Vision Language Models (VLMs) and OCR engines on structural layout analysis, font recognition, and document deg

Models trained on arabic-ocr-synthetic-scans-faker-300k

2 models list it as training data.

Open loay/arabic-ocr-synthetic-scans-faker-300k on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.