Sprakbanken/synthetic_sami_ocr_data download history

Sprakbanken/synthetic_sami_ocr_data is an image to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 117 times (37 in the last 7 days), and 8,675 times in total. It ranks #101,856 among datasets by monthly downloads.

Synthetic text images for North, South, Lule and Inari Sámi This dataset contains synthetic line images meant for fitting OCR models for North, South, Lule and Inari Sámi. Clean line images are created using Pillow and they are subsequently distorted using Augraphy [1]. Text sourc

Models trained on synthetic_sami_ocr_data

3 models list it as training data.

Open Sprakbanken/synthetic_sami_ocr_data on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.