Load 4bit models 4x faster Collection Native bitsandbytes 4bit pre quantized models • 25 items • Updated 3 days ago • 62
TinyLlama/TinyLlama-1.1B-Chat-v1.0 Text Generation • 1B • Updated Mar 17, 2024 • 2.48M • • 1.72k
CohereLabs/c4ai-command-r-plus-4bit Text Generation • 105B • Updated Apr 16, 2025 • 727 • 262
ContextualAI/Contextual_KTO_Mistral_PairRM Text Generation • 7B • Updated Apr 26, 2024 • 63 • 33