YuehHanChen/SAGE-Eval download history
YuehHanChen/SAGE-Eval is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 31 times (3 in the last 7 days), and 436 times in total. It ranks #266,938 among datasets by monthly downloads.
[NeurIPS 2025 Spotlight] SAGE‑Eval — Safety‑fact Systematic Generalization Benchmark by Yueh‑Han Chen, Guy Davidson, Brenden M. Lake Correspondence to yc7592@nyu.edu SAGE‑Eval, SAfety-fact systematic GEneralization evaluation, is a benchmark for evaluating whether large language models
Open YuehHanChen/SAGE-Eval on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.