amd/Llama-2-70b-chat-hf-WMXFP4-AMXFP4-KVFP8-Scale-UINT8-MLPerf-GPTQ download history

amd/Llama-2-70b-chat-hf-WMXFP4-AMXFP4-KVFP8-Scale-UINT8-MLPerf-GPTQ is a 36.9B-parameter model model by amd. In the last 30 days it was downloaded 16 times (7 in the last 7 days), and 1,465 times in total.

It ranks #502,791 on the Hub by monthly downloads.

It has 0 likes.

It is a quantization of meta-llama/Llama-2-70b-chat-hf.

Open amd/Llama-2-70b-chat-hf-WMXFP4-AMXFP4-KVFP8-Scale-UINT8-MLPerf-GPTQ on Hugging Face