llama based models quantized to various precisions
-
jsbaicenter/Llama-3.3-70B-Instruct-FP8-Dynamic
Text Generation • 71B • Updated • 23 -
jsbaicenter/Llama-3.2-1b-Instruct-AWQ-4bit-GEMM
1B • Updated • 6 -
jsbaicenter/Llama-3.2-1B-Instruct-FP8
Text Generation • 1B • Updated • 11 -
jsbaicenter/Llama-3.2-3b-Instruct-AWQ-4bit-GEMM
Text Generation • 3B • Updated • 11