comma-project/comma-jsonl download history
comma-project/comma-jsonl is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 68 times (12 in the last 7 days), and 1,686 times in total. It ranks #149,549 among datasets by monthly downloads.
Dataset Card for CoMMA JSON-L CoMMA is a large-scale corpus of digitized medieval manuscripts transcribed using Handwritten Text Recognition (HTR). It contains over 2.5 billion tokens from more than 23,000 manuscripts in Latin and Old French (801–1600 CE). Unlike most existing resources
Models trained on comma-jsonl
1 models list it as training data.
- comma-project/modernbert 236 downloads in 30 days