How to use from
Lemonade
Pull the model
# Download Lemonade from https://lemonade-server.ai/
lemonade pull aashishk029/shinobu-7b-v13:Q4_K_M
Run and chat with the model
lemonade run user.shinobu-7b-v13-Q4_K_M
List all available models
lemonade list
Quick Links

Shinobu 7B v13 β€” on-prem security triage copilot

Qwen2.5-7B-Instruct + Shinobu LoRA v13, Q4_K_M GGUF (4.4 GB). The L2 triage brain of Shinobu β€” a privacy-first, 100% local security copilot. Detects, triages (MITRE ATT&CK-mapped verdicts with abstain), and recommends operator-approved responses. Nothing auto-executes.

  • 87% verdict accuracy on probes_v11 (TP 9/10, FP 9/10, abstain 8/10) β€” up from v10's 80%
  • Abstains rather than fabricating verdicts on thin telemetry
  • sha256: cf493ea6178a52802c18796d41be71b52e7d8ca8c0b48fe14fed1fd3099cf75c
  • Runs via ollama on 16 GB RAM / 8-core CPU, no GPU

Install + full product: https://github.com/aashishk029/shinobu β€” created by Aashish Kumar.

Downloads last month
14
GGUF
Model size
8B params
Architecture
qwen2
Hardware compatibility
Log In to add your hardware

4-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for aashishk029/shinobu-7b-v13

Base model

Qwen/Qwen2.5-7B
Adapter
(2564)
this model