Sprakbanken/synthetic_sami_ocr_data download history
Sprakbanken/synthetic_sami_ocr_data is an image to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 117 times (37 in the last 7 days), and 8,675 times in total. It ranks #101,856 among datasets by monthly downloads.
Synthetic text images for North, South, Lule and Inari Sámi This dataset contains synthetic line images meant for fitting OCR models for North, South, Lule and Inari Sámi. Clean line images are created using Pillow and they are subsequently distorted using Augraphy [1]. Text sourc
Models trained on synthetic_sami_ocr_data
3 models list it as training data.
- Sprakbanken/trocr_smi_nor_pred_synth 23 downloads in 30 days
- Sprakbanken/trocr_smi_synth 21 downloads in 30 days
- Sprakbanken/trocr_smi_pred_synth 20 downloads in 30 days
Open Sprakbanken/synthetic_sami_ocr_data on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.