Quant generated from: https://huggingface.co/Orenguteng/Llama-3.1-8B-Lexi-Uncensored-V2

The original GGUF quants provided by Orenguteng (https://huggingface.co/Orenguteng/Llama-3.1-8B-Lexi-Uncensored-V2-GGUF) did not include a Q6_K, thus I generated one.

At 6.6 GB the Q6_K is a great high quality option for 8GB GPUs.

Downloads last month
427
GGUF
Model size
8B params
Architecture
llama
Hardware compatibility
Log In to add your hardware

6-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for Spaceballs/Llama-3.1-8B-Lexi-Uncensored-V2-Q6_K-GGUF

Quantized
(32)
this model

Collection including Spaceballs/Llama-3.1-8B-Lexi-Uncensored-V2-Q6_K-GGUF