Inference Providers
Active filters: w4a16
ramblingpolymath/Qwen3-Coder-30B-A3B-Instruct-W4A16
Text Generation
• 31B • Updated • 26
• 2
ramblingpolymath/Qwen3-30B-A3B-Thinking-2507-W4A16
Text Generation
• 31B • Updated • 8
• 3
twhitworth/gpt-oss-120b-awq-w4a16
117B • Updated • 3.38k
• 24
TheHouseOfTheDude/Behemoth-R1-123B-v2_Compressed-Tensors
Text Generation
• Updated • 2
TheHouseOfTheDude/Behemoth-X-123B-v2_Compressed-Tensors
Text Generation
• Updated • 1
TheHouseOfTheDude/GLM-Steam-106B-A12B-v1_Compressed-Tensors
Text Generation
• Updated TheHouseOfTheDude/L3.3-Animus-V10.0_Compressed-Tensors
Text Generation
• Updated TheHouseOfTheDude/Behemoth-ReduX-123B-v1_Compressed-Tensors
Text Generation
• Updated • 3
TheHouseOfTheDude/Qwen3-Next-80B-A3B-Instruct_Compressed-Tensors
Text Generation
• Updated • 8
TheHouseOfTheDude/Fallen-Command-A-111B-v1_Compresses-Tensors
Text Generation
• Updated TheHouseOfTheDude/Behemoth-ReduX-123B-v1.1_Compressed-Tensors
Text Generation
• Updated • 4
TheHouseOfTheDude/L3.3-70B-Animus-V12.0_Compressed-Tensors
Text Generation
• Updated TheHouseOfTheDude/Behemoth-X-123B-v2.1_Compressed-Tensors
Text Generation
• Updated • 1
ModelCloud/GLM-4.6-GPTQMODEL-W4A16-v1
Text Generation
• 357B • Updated • 9
ModelCloud/GLM-4.6-GPTQMODEL-W4A16-v2
Text Generation
• 357B • Updated • 10
• 1
ModelCloud/GLM-4.6-REAP-268B-A32B-GPTQMODEL-W4A16
Text Generation
• 269B • Updated • 5
• 2
ModelCloud/MiniMax-M2-GPTQMODEL-W4A16
Text Generation
• 229B • Updated • 17
• 3
ModelCloud/Marin-32B-Base-GPTQMODEL-W4A16
Text Generation
• 33B • Updated • 5
• 1
ModelCloud/Marin-32B-Base-GPTQMODEL-AWQ-W4A16
Text Generation
• 33B • Updated • 13
• 2
TheHouseOfTheDude/Legion-V2.1-LLaMa-70B_CompressedTensors
Text Generation
• Updated ModelCloud/Granite-4.0-H-1B-GPTQMODEL-W4A16
Text Generation
• 1B • Updated • 6
• 1
ModelCloud/Granite-4.0-H-350M-GPTQMODEL-W4A16
Text Generation
• 0.3B • Updated • 5
• 1
ModelCloud/Brumby-14B-Base-GPTQMODEL-W4A16
Text Generation
• 15B • Updated • 7
• 1
ModelCloud/Brumby-14B-Base-GPTQMODEL-W4A16-v2
Text Generation
• 15B • Updated • 6
• 1
TheHouseOfTheDude/Precog-24B-v1_Compressed-Tensors
Text Generation
• Updated TheHouseOfTheDude/M2411-123B-Animus-V12.0_Compressed-Tensors
Text Generation
• Updated • 2
TheHouseOfTheDude/INTELLECT-3_Compressed-Tensors
Text Generation
• Updated • 1
avtc/GLM-4.6-REAP-268B-A32B-GPTQMODEL-W4A16
Text Generation
• 271B • Updated • 40
• 3
avtc/MiniMax-M2-GPTQMODEL-W4A16
Text Generation
• 229B • Updated • 8
0xSero/GLM-4.6-218B-W4A16
Text Generation
• 2B • Updated • 12
• 8