fatihburakkaragoz/anadolu-ocr-corpus download history
fatihburakkaragoz/anadolu-ocr-corpus is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 24 times (8 in the last 7 days), and 500 times in total. It ranks #327,226 among datasets by monthly downloads.
Anadolu OCR Corpus Anadolu OCR Corpus is an OpenCR export of OCR text and document metadata for 52 historical Ottoman Turkish, Turkish, and Arabic-containing PDF sources. The dataset is provided in two Hugging Face configs: pages: one row per source page, including page-level OCR text, m
Models trained on anadolu-ocr-corpus
1 models list it as training data.
- fatihburakkaragoz/fuzuli-base 0 downloads in 30 days
Open fatihburakkaragoz/anadolu-ocr-corpus on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.