formalmathatepfl/qwen3-sft-feedback-with-proof-repair-rl-125-unmasked download history
formalmathatepfl/qwen3-sft-feedback-with-proof-repair-rl-125-unmasked is an 8.2B-parameter model model by formalmathatepfl. In the last 30 days it was downloaded 54 times (3 in the last 7 days), and 54 times in total.
It ranks #245,922 on the Hub by monthly downloads.
It has 0 likes.
1 models build on qwen3-sft-feedback-with-proof-repair-rl-125-unmasked: 0 quantized, 1 fine-tuned, 0 adapters and 0 merges. Together with the original they were downloaded 1,122 times in the last 30 days. See the qwen3-sft-feedback-with-proof-repair-rl-125-unmasked galaxy.
Most downloaded derivatives
- formalmathatepfl/feedback-grpo-reasoning-sft (finetune), 1.1K downloads in 30 days
Open formalmathatepfl/qwen3-sft-feedback-with-proof-repair-rl-125-unmasked on Hugging Face