hmar-heritage-org/paragraphs download history

hmar-heritage-org/paragraphs is a fill mask dataset on the Hugging Face Hub. In the last 30 days it was downloaded 36 times (6 in the last 7 days), and 36 times in total. It ranks #237,432 among datasets by monthly downloads.

paragraphs A multi-register, long-context text corpus for the Hmar language (hmr, ISO 639-3), part of the Zo languages family. It contains 44,162 training sequences and 941 evaluation sequences (~4.37 million words), formatted into complete paragraphs up to 512 tokens. Maintained by t

Models trained on paragraphs

2 models list it as training data.

Spaces using paragraphs

Open hmar-heritage-org/paragraphs on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.