spadeMIA/GoodWiki_Corpus_1024_2040 download history

spadeMIA/GoodWiki_Corpus_1024_2040 is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 221 times (34 in the last 7 days), and 604 times in total. It ranks #62,310 among datasets by monthly downloads.

GoodWiki 1024–2040: paragraph-truncated MIA fine-tuning corpus A deterministic, paragraph-truncated corpus of English Wikipedia Good/Featured articles, built from euirim/goodwiki for membership inference attack (MIA) experiments on fine-tuned language models. Membership labels are de

Models trained on GoodWiki_Corpus_1024_2040

3 models list it as training data.

Open spadeMIA/GoodWiki_Corpus_1024_2040 on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.