How to use from
Hermes Agent
Start the llama.cpp server
# Install llama.cpp:
brew install llama.cpp
# Start a local OpenAI-compatible server:
llama serve -hf cumhuronat/OnatSec-27B:BF16
Configure Hermes
# Install Hermes:
curl -fsSL https://hermes-agent.nousresearch.com/install.sh | bash
hermes setup
# Point Hermes at the local server:
hermes config set model.provider custom
hermes config set model.base_url http://127.0.0.1:8080/v1
hermes config set model.default cumhuronat/OnatSec-27B:BF16
Run Hermes
hermes
Quick Links

OnatSec-27B-v1.1

Fine-tune of Qwen/Qwen3.6-27B targeting CS-Eval โ€” currently the top-performing open-weight model on the leaderboard.

The Multi-Token Prediction (MTP) head is preserved through conversion and quantization (blk.64.* kept at Q8_0), so self-speculative decoding works out of the box for a ~1.5โ€“2x decode speedup.

Downloads last month
457
GGUF
Model size
27B params
Architecture
qwen35
Hardware compatibility
Log In to add your hardware

4-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. ๐Ÿ™‹ Ask for provider support

Model tree for cumhuronat/OnatSec-27B

Base model

Qwen/Qwen3.6-27B
Quantized
(679)
this model