Inference Providers
Active filters: edge-ai
magicunicorn/whisper-large-v3-amd-npu-int8
Updated • 6
• 4
magicunicorn/whisper-large-v2-amd-npu-int8
Updated • 3
• 5
magicunicorn/whisper-medium-amd-npu-int8
magicunicorn/whisper-small-amd-npu-int8
Updated • 4
• 1
magicunicorn/whisper-base-amd-npu-int8
jamescallander/deepseek-llm-7b-chat_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 39
jamescallander/txgemma-9b-chat_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 4
jamescallander/NextCoder-7B_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 25
• 1
jamescallander/Qwen2.5-Math-1.5B-Instruct_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 32
jamescallander/MediPhi-Instruct_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 8
• 1
jamescallander/Qwen2.5-Coder-3B-Instruct_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 124
• 2
jamescallander/Llama-3.1-8B-Instruct_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 35
jamescallander/MiniCPM4-0.5B_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 25
jamescallander/TinyLlama-1.1B-Chat-v1.0_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 23
• 2
jamescallander/MediPhi-PubMed_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 5
• 1
jamescallander/medgemma-4b-it_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 1
• 1
jamescallander/Qwen2.5-Math-7B-Instruct_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 10
jamescallander/deepseek-coder-1.3b-instruct_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 30
jamescallander/MiniCPM3-4B_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 22
• 1
jamescallander/Llama-3.2-3B-Instruct_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 92
cluangar/typhoon2-qwen2.5-7B-rk3588
Text Generation
• Updated • 1
• 1
jamescallander/DeepSeek-R1-Distill-Qwen-14B_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 68.2k
• 1
jamescallander/deepseek-coder-6.7b-instruct_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 4
jamescallander/CodeLlama-7b-Instruct-hf_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 9
Text Generation
• 0.8B • Updated • 3.67k
• 1
jamescallander/DeepSeek-R1-Distill-Qwen-1.5B_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 176
jamescallander/DeepSeek-R1-Distill-Llama-8B_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 20
Zhare-AI/janus-pro-7b-webgpu
Image-to-Text
• Updated • 31
• 2
ForeseeLab/foreseeai-qwen3-4b-iot-int4
Text Generation
• 4B • Updated • 8
• 1
jamescallander/gemma-2-2b-it_w8a8_g128_rk3588.rkllm
Text Generation
• Updated • 27