How to use from
Lemonade
Pull the model
# Download Lemonade from https://lemonade-server.ai/
lemonade pull ggml-org/Qwen3-8B-Base-GGUF:BF16
Run and chat with the model
lemonade run user.Qwen3-8B-Base-GGUF-BF16
List all available models
lemonade list
Quick Links

Qwen3-8B-Base

Run with https://llama.app

llama serve -hf ggml-org/Qwen3-8B-Base-GGUF

Source models

This model is automatically converted using https://github.com/ggml-org/convert

Downloads last month
308
GGUF
Model size
8B params
Architecture
qwen3
Hardware compatibility
Log In to add your hardware

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for ggml-org/Qwen3-8B-Base-GGUF

Quantized
(46)
this model