kensho/WILD download history
kensho/WILD is a text classification dataset on the Hugging Face Hub. In the last 30 days it was downloaded 152 times (43 in the last 7 days), and 1,022 times in total. It ranks #83,435 among datasets by monthly downloads.
Dataset to accompany the paper Cost-Efficient Estimation of General Abilities Across Benchmarks. WILD: Wide-scale Item Level Dataset WILD is a large-scale evaluation response matrix containing item-level binary scores for 65 language models across 27 benchmarks (163 subtasks, 109,566 uniq