njand/latin-asr-post-processing-dataset download history

njand/latin-asr-post-processing-dataset is a token classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 24 times (2 in the last 7 days), and 142 times in total. It ranks #326,781 among datasets by monthly downloads.

Latin ASR Post-Processing Dataset A sequence-labeling dataset built for fine-tuning BERT-style models (e.g., latin-bert) on Inverse Text Normalization (ITN) — restoring capitalization and punctuation on raw, lowercased Latin text (such as njand/wav2vec2-xls-r-latin ASR outputs). Compi

Models trained on latin-asr-post-processing-dataset

1 models list it as training data.

Open njand/latin-asr-post-processing-dataset on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.