kazuyamaa/Qwen3-4B-PPO-3000data-v1 download history
kazuyamaa/Qwen3-4B-PPO-3000data-v1 is a reinforcement learning model by kazuyamaa. In the last 30 days it was downloaded 7 times (2 in the last 7 days), and 34 times in total.
It ranks #1,056,990 on the Hub by monthly downloads and #21,539 among reinforcement learning models.
It has 0 likes.
Open kazuyamaa/Qwen3-4B-PPO-3000data-v1 on Hugging Face
Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.