translategemma-4b-it-Q2_K-GGUF

GGUF-quantized derivative of Google's TranslateGemma 4B for on-device translation.

Base Model

Conversion Details

  • Source GGUF: mradermacher/translategemma-4b-it-GGUF
  • Quantization: Q2_K
  • GGUF Size: ~1.7 GB
  • Prepared by: Zanish Labs / Voco

Runtime

  • Compatible with llama.cpp (CPU/NEON)
  • Tested on Voco iOS app (STQ1_0 backend, PR #22836)

Attribution

This is a quantized derivative. The original model was created by Google. See LICENSE for full license text.

Links

Downloads last month
16
GGUF
Model size
4B params
Architecture
gemma3
Hardware compatibility
Log In to add your hardware

2-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for zanish-labs/translategemma-4b-it-Q2_K-gguf

Quantized
(39)
this model