inference-optimization/Qwen3-8B-DFlash-FP8-BLOCK download history

inference-optimization/Qwen3-8B-DFlash-FP8-BLOCK is a 2.3B-parameter text generation model by inference-optimization. In the last 30 days it was downloaded 31 times (31 in the last 7 days), and 31 times in total.

It ranks #310,277 on the Hub by monthly downloads and #84,715 among text generation models.

It has 0 likes.

It is a quantization of Qwen/Qwen3-8B.

Open inference-optimization/Qwen3-8B-DFlash-FP8-BLOCK on Hugging Face