weixu-zhang/visrl-bench-results download history

weixu-zhang/visrl-bench-results is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 5,978 times (5,911 in the last 7 days), and 6,934 times in total. It ranks #4,882 among datasets by monthly downloads.

VisRL benchmark eval results — VisPlotBench per-sample outputs Companion to github.com/weixuzhang/visrl and weixu-zhang/visrl-data. This repo holds the per-sample outputs of VisPlotBench runs for 5 trained checkpoints, so the GPT-4.1 visual judge or the custom HTML/vegalite/lilypond s

Open weixu-zhang/visrl-bench-results on Hugging Face