hmar-heritage-org/paragraphs download history
hmar-heritage-org/paragraphs is a fill mask dataset on the Hugging Face Hub. In the last 30 days it was downloaded 36 times (6 in the last 7 days), and 36 times in total. It ranks #237,432 among datasets by monthly downloads.
paragraphs A multi-register, long-context text corpus for the Hmar language (hmr, ISO 639-3), part of the Zo languages family. It contains 44,162 training sequences and 941 evaluation sequences (~4.37 million words), formatted into complete paragraphs up to 512 tokens. Maintained by t
Models trained on paragraphs
2 models list it as training data.
- azinamotoe/HmarBERT-mini 410 downloads in 30 days
- azinamotoe/HmarBERT-mini-v2 34 downloads in 30 days
Spaces using paragraphs
- HmarBERT Arena 4 likes
Open hmar-heritage-org/paragraphs on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.