llama.cpp b9637 Windows runtime mirror

This repository mirrors an immutable subset of the official ggml-org/llama.cpp b9637 release for resilient application installation on Windows x64.

Included archives:

  • Windows x64 CPU
  • Windows x64 Vulkan
  • Windows x64 CUDA 12.4 executable/runtime files built by llama.cpp
  • Windows x64 CUDA 13.3 executable/runtime files built by llama.cpp

Every file is unmodified. Verify it against checksums.sha256 before use.

NVIDIA CUDA runtime and cuBLAS redistributable archives are deliberately not republished here. Applications that select a CUDA build must obtain those components from the official upstream release under the NVIDIA CUDA EULA, or fall back to the Vulkan build.

This mirror is not affiliated with or endorsed by the llama.cpp project or NVIDIA.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support