joelbarmettler/gheim-ch-pii-212k download history
joelbarmettler/gheim-ch-pii-212k is a token classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 25 times (8 in the last 7 days), and 225 times in total. It ranks #316,899 among datasets by monthly downloads.
gheim-ch-pii-212k Summary. 212,503-chunk multilingual PII NER dataset covering the four official Swiss languages and English. 84% is real text from the Apertus pretrain corpora (Swiss court rulings, federal parliament records, Swiss-filtered web text, Romansh corpus); the remaining
Models trained on gheim-ch-pii-212k
2 models list it as training data.
- joelbarmettler/gheim-ch-560m 170 downloads in 30 days
- joelbarmettler/gheim-ch-560m-research 38 downloads in 30 days
Open joelbarmettler/gheim-ch-pii-212k on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.