bigdatamark/synthid-text-research download history

bigdatamark/synthid-text-research is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 189 times (122 in the last 7 days), and 189 times in total. It ranks #70,663 among datasets by monthly downloads.

Corpus provenance Every corpus used in the SynthID detection analysis, with exactly what produced it. All corpora use collection_prompt.txt verbatim (13 unnamed four-token referents, 800-word target). Vendor corpora (UNKNOWN watermark status — these are the test subjects)

Open bigdatamark/synthid-text-research on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.