himalaya-ai/nepali-ocr-corpus download history

himalaya-ai/nepali-ocr-corpus is an image to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 450 times (450 in the last 7 days), and 450 times in total. It ranks #35,993 among datasets by monthly downloads.

nepali-ocr-corpus A fine-tuning corpus for Nepali (Devanagari) government-document OCR and layout analysis, built to adapt a vision-language model (dots.ocr) to Nepali paperwork without erasing what it already knows. About half the corpus by tokens is the target task: synthetic and re

Open himalaya-ai/nepali-ocr-corpus on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.