amishor/reinforce-learning download history

amishor/reinforce-learning is a dataset on the Hugging Face Hub. In the last 30 days it was downloaded 44 times (7 in the last 7 days), and 571 times in total. It ranks #203,866 among datasets by monthly downloads.

DAPO-RL-Instruct Dataset A high-quality instruction-following dataset derived from the open-source technical report “DAPO: An Open-Source LLM Reinforcement Learning System at Scale” (arXiv:2503.14476, March 2025). This dataset captures key concepts, training strategies, and system design

Open amishor/reinforce-learning on Hugging Face