LeonOverload/primo-rl-json download history

LeonOverload/primo-rl-json is a video text to text dataset on the Hugging Face Hub. In the last 30 days it was downloaded 234 times (55 in the last 7 days), and 967 times in total. It ranks #59,859 among datasets by monthly downloads.

PRIMO RL Data Stage-2 (GRPO reinforcement learning) training annotations for PRIMO R1 (paper). Unlike the SFT data, these records carry no chain-of-thought traces — RL optimizes against a verifiable progress reward, so only the ground-truth answer is needed. That is the point of the m

Open LeonOverload/primo-rl-json on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.