ngqtrung/omr-grpo-train download history
ngqtrung/omr-grpo-train is an image text to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 11 times (1 in the last 7 days), and 172 times in total. It ranks #540,266 among datasets by monthly downloads.
OMR GRPO Train — Math-Image RL Training Set 74,971 rows · 8 sources · images embedded as bytes (self-contained) This is the RL training parquet used for Group Relative Policy Optimization (GRPO) on math and STEM image-reasoning tasks. It was built by merging and filtering eight public
Open ngqtrung/omr-grpo-train on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.