library_name: gguf
tags:
- GGUF
- Llama
- quantization
base_model: - Orenguteng/Llama-3.1-8B-Lexi-Uncensored-V2
license: llama3.1
pipeline_tag: text-generation
Quant generated from: https://huggingface.co/Orenguteng/Llama-3.1-8B-Lexi-Uncensored-V2
The original GGUF quants provided by Orenguteng (https://huggingface.co/Orenguteng/Llama-3.1-8B-Lexi-Uncensored-V2-GGUF) did not include a Q6_K, thus I generated one.
At 6.6 GB the Q6_K is a great high quality option for 8GB GPUs.