toiar/Khasi-OCR-21K download history
toiar/Khasi-OCR-21K is an image to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 8 times (3 in the last 7 days), and 67 times in total. It ranks #638,775 among datasets by monthly downloads.
Khasi-OCR-21K Khasi-OCR-21K is a curated Vision-Language dataset totaling 21,319 samples, specifically designed to train robust OCR models for the Khasi language. This version introduces a significant amount of high-quality real book data alongside synthetic samples to handle diverse docu
Open toiar/Khasi-OCR-21K on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.