Nellyw888/VeriReason-Qwen2.5-1.5b-RTLCoder-Verilog-GRPO-reasoning-tb download history

Nellyw888/VeriReason-Qwen2.5-1.5b-RTLCoder-Verilog-GRPO-reasoning-tb is a 1.5B-parameter reinforcement learning model by Nellyw888. In the last 30 days it was downloaded 23 times (4 in the last 7 days), and 410 times in total.

It ranks #371,695 on the Hub by monthly downloads and #7,322 among reinforcement learning models.

It has 1 likes.

It is a fine-tune of Qwen/Qwen2.5-Coder-1.5B-Instruct.

2 models build on VeriReason-Qwen2.5-1.5b-RTLCoder-Verilog-GRPO-reasoning-tb: 2 quantized, 0 fine-tuned, 0 adapters and 0 merges. Together with the original they were downloaded 979 times in the last 30 days. See the VeriReason-Qwen2.5-1.5b-RTLCoder-Verilog-GRPO-reasoning-tb galaxy.

Most downloaded derivatives

Open Nellyw888/VeriReason-Qwen2.5-1.5b-RTLCoder-Verilog-GRPO-reasoning-tb on Hugging Face

Sister project: Paper Pulse, the upvote history of every Hugging Face Daily Paper.