joelbarmettler/gheim-ch-pii-212k download history

joelbarmettler/gheim-ch-pii-212k is a token classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 25 times (8 in the last 7 days), and 225 times in total. It ranks #316,899 among datasets by monthly downloads.

gheim-ch-pii-212k Summary. 212,503-chunk multilingual PII NER dataset covering the four official Swiss languages and English. 84% is real text from the Apertus pretrain corpora (Swiss court rulings, federal parliament records, Swiss-filtered web text, Romansh corpus); the remaining

Models trained on gheim-ch-pii-212k

2 models list it as training data.

Open joelbarmettler/gheim-ch-pii-212k on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.