base_model: Rangle2/gemma-4-12B-it-uncensored-opus4.7-cot
base_model_relation: quantized
library_name: gguf
pipeline_tag: text-generation
tags:
- gguf
- llama.cpp
- quantized
- gemma
gemma-4-12B-it-uncensored-opus4.7-cot — GGUF
GGUF quantizations of Rangle2/gemma-4-12B-it-uncensored-opus4.7-cot.
Produced with llama.cpp (llama-quantize) from an F16 GGUF converted directly from the base safetensors.
Available quants
| File | Quant | Size |
|---|---|---|
gemma-4-12B-it-uncensored-opus4.7-cot-Q2_K.gguf |
Q2_K | 4.83 GB |
gemma-4-12B-it-uncensored-opus4.7-cot-Q3_K_M.gguf |
Q3_K_M | 6.09 GB |
gemma-4-12B-it-uncensored-opus4.7-cot-Q4_K_M.gguf |
Q4_K_M | 7.38 GB |
gemma-4-12B-it-uncensored-opus4.7-cot-Q5_K_M.gguf |
Q5_K_M | 8.55 GB |
gemma-4-12B-it-uncensored-opus4.7-cot-Q6_K.gguf |
Q6_K | 9.79 GB |
gemma-4-12B-it-uncensored-opus4.7-cot-Q8_0.gguf |
Q8_0 | 12.67 GB |
gemma-4-12B-it-uncensored-opus4.7-cot-F16.gguf |
F16 | 23.83 GB |
Usage
llama-cli -hf Rangle2/gemma-4-12B-it-uncensored-opus4.7-cot-GGUF:Q4_K_M