WardenSight-1B — GGUF

Official GGUF builds of WardenSight-1B, converted by the model author from the original safetensors (SHA-256 070345c0…11f16) with llama.cpp's converter.

WardenSight-1B is a 1-billion parameter terminal agent specialized for vulnerability file localization in WordPress plugins: given a plugin repository and a CWE class, it explores the code through a read-only terminal interface and submits a ranked list of files it believes contain that weakness. It is a triage instrument — it tells an analyst which files to read first; it does not produce line numbers, exploits, patches, or proof that code is vulnerable.

See the main model card for the full evaluation (File-F1 0.3263 vs base 0.2671 on 1,646 held-out WordPress plugin tasks, +22.2% relative), intended use, limitations, and the agent prompt contract.

Files

File Quant Size Notes
wardensight-1b.Q4_K_M.gguf Q4_K_M 1.02 GB Recommended default — best size/quality balance
wardensight-1b.Q8_0.gguf Q8_0 1.74 GB Near-lossless
wardensight-1b.BF16.gguf BF16 3.27 GB Full precision; use as a source for re-quantizing

Usage

Runs anywhere llama.cpp does, including Apple Metal:

llama-server -m wardensight-1b.Q4_K_M.gguf -c 16384

Or load directly in LM Studio / Ollama. WardenSight is an agentic model: it expects the terminal-agent prompt contract described in the main model card and performs best driven by the Antares CLI harness documented there.

License

Apache-2.0, inherited unchanged from the safetensors release.

Downloads last month
195
GGUF
Model size
2B params
Architecture
granite
Hardware compatibility
Log In to add your hardware

4-bit

8-bit

16-bit

Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for emiluzelac/wardensight-1b-GGUF

Quantized
(1)
this model