willamazon1/sdft-search-lora-iter160 download history
willamazon1/sdft-search-lora-iter160 is an 8.2B-parameter reinforcement learning model by willamazon1. In the last 30 days it was downloaded 8 times (2 in the last 7 days), and 30 times in total.
It ranks #932,135 on the Hub by monthly downloads and #22,273 among reinforcement learning models.
It has 0 likes.
It is an adapter of Qwen/Qwen3-8B-Base.
1 models build on sdft-search-lora-iter160: 0 quantized, 0 fine-tuned, 0 adapters and 1 merges. Together with the original they were downloaded 263 times in the last 30 days. See the sdft-search-lora-iter160 galaxy.
Most downloaded derivatives
- willhx/Qwen3-8B-SDFT-MLE-Math-Search-Ecom (merge), 255 downloads in 30 days