thekamilya/kazakh-printed-dataset download history

thekamilya/kazakh-printed-dataset is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 20 times (6 in the last 7 days), and 162 times in total. It ranks #374,093 among datasets by monthly downloads.

Kazakh Printed Dataset for OCR task Data Lineage This dataset was synthetically generated using issai/kazparc as the base. Since Kazakh OCR data is scarce, I developed a pipeline to transform digital Kazakh text into a printed-style dataset. Generation Process

Models trained on kazakh-printed-dataset

1 models list it as training data.

Open thekamilya/kazakh-printed-dataset on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.