inference-optimization/Qwen3-8B-DFlash-GPTQ-IMatrix-Gauss-NVFP4-W4A4 download history
inference-optimization/Qwen3-8B-DFlash-GPTQ-IMatrix-Gauss-NVFP4-W4A4 is a 1.0B-parameter text generation model by inference-optimization. In the last 30 days it was downloaded 34 times (34 in the last 7 days), and 34 times in total.
It ranks #297,040 on the Hub by monthly downloads and #80,472 among text generation models.
It has 0 likes.
It is a quantization of Qwen/Qwen3-8B.
Open inference-optimization/Qwen3-8B-DFlash-GPTQ-IMatrix-Gauss-NVFP4-W4A4 on Hugging Face