foundation-models/eval-runs download history

foundation-models/eval-runs is a text generation dataset on the Hugging Face Hub. In the last 30 days it was downloaded 9 times (3 in the last 7 days), and 108 times in total. It ranks #601,102 among datasets by monthly downloads.

eval-runs Evaluation run artifacts from τ2-bench simulations. Layout tau2-bench/ gpt-4o-mini/ with-patch/ # Retail policy includes agent-lens write guardrail without-patch/ # Baseline τ2-bench retail policy (no guardrail) other-models/ with-patc

Open foundation-models/eval-runs on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.