Interplay-LM-Reasoning/composition download history
Interplay-LM-Reasoning/composition is a question answering dataset on the Hugging Face Hub. In the last 30 days it was downloaded 463 times (121 in the last 7 days), and 6,835 times in total. It ranks #35,191 among datasets by monthly downloads.
On the Interplay of Pre-Training, Mid-Training, and RL on Reasoning Language Models Charlie Zhang, Graham Neubig, Xiang Yue Carnegie Mellon University, Language Technologies Institute Does Reinforcement Learning Truly Extend Reasoning? This work explores the discrepancy in vi
Open Interplay-LM-Reasoning/composition on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.