Making a 27B reasoning-capable model practical for long-horizon agentic workflows and specialist cyber evaluations is the core trade-off this release targets: it compresses a full Qwen3.8-27B checkpoint into a GPU-friendly GGUF quant while intentionally exposing an "uncensored" behavior profile for interpretability and red-teaming.
Key Capabilities
- Extreme compression with high fidelity: OrcaSAQ2-style quantization reduces the original BF16 checkpoint (54 GB) to roughly 12.3 GB while reporting only +0.02% perplexity, ~93.2% token-level top-1 agreement, and mean KLD ≈ 0.031 on calibration tests — preserving long-context (262k) and reasoning behavior.
- Architecture and feature preservation: the hybrid Gated DeltaNet + full-attention architecture, the MTP speculative-decoding head, thinking mode and tool-calling capabilities are retained; vision support is provided via a separate mmproj file and the build targets llama.cpp / GGUF runtimes.
- Research and red-team orientation: the build is explicitly "abliterated" (refusal directions suppressed) to study refusal mechanisms, robustness, and offensive-security capabilities; users are expected to add their own safety and moderation layers before any deployment.
Who it's for and tradeoffs
Great fit if you are doing offline research, interpretability or red-teaming on large multimodal LLMs and need a compact, long-context-capable 27B that still supports thinking/tooling primitives. Look elsewhere if you need a production-safe model out-of-the-box: the uncensored nature removes many safety refusals and increases misuse risk. Additional tradeoffs include modest fidelity degradation at extreme low-bit quants and the requirement of a recent llama.cpp build (qwen35 + nextn/MTP support) or compatible runtimes to fully restore architecture and speculative decoding behavior.
In short: a practical, high-fidelity quantized path to run a reasoning-capable Qwen3.8-27B variant for long-horizon agents and cyber-focused research, but not a turnkey solution for public-facing production without substantial safety controls.