himalaya-ai/nepali-ocr-corpus download history
himalaya-ai/nepali-ocr-corpus is an image to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 450 times (450 in the last 7 days), and 450 times in total. It ranks #35,993 among datasets by monthly downloads.
nepali-ocr-corpus A fine-tuning corpus for Nepali (Devanagari) government-document OCR and layout analysis, built to adapt a vision-language model (dots.ocr) to Nepali paperwork without erasing what it already knows. About half the corpus by tokens is the target task: synthetic and re
Open himalaya-ai/nepali-ocr-corpus on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.