comma-project/comma-jsonl download history

comma-project/comma-jsonl is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 68 times (12 in the last 7 days), and 1,686 times in total. It ranks #149,549 among datasets by monthly downloads.

Dataset Card for CoMMA JSON-L CoMMA is a large-scale corpus of digitized medieval manuscripts transcribed using Handwritten Text Recognition (HTR). It contains over 2.5 billion tokens from more than 23,000 manuscripts in Latin and Old French (801–1600 CE). Unlike most existing resources

Models trained on comma-jsonl

1 models list it as training data.

Open comma-project/comma-jsonl on Hugging Face