nevmenandr/russian-old-orthography-ocr download history
nevmenandr/russian-old-orthography-ocr is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 777 times (227 in the last 7 days), and 13,631 times in total. It ranks #23,444 among datasets by monthly downloads.
Basic Description Dataset contains source images and human-readable extracted texts. All texts were published in Russia in the 19th century and written using pre-reform orthography. The dataset is designed to train and evaluate optical character recognition systems for texts published in
Open nevmenandr/russian-old-orthography-ocr on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.