AIAny
AI Model2026
Icon for item

Qwen3.8-27B — OBLITERATED

An uncensored, weight-modified variant of Qwen3.8-27B that surgically removes the model's refusal directions to produce 0% refusals while aiming to preserve or improve capability. Uses complementary abliteration blending (SVD + LEACE blend) and ships with recommended greedy inference settings; intended for AI-safety research and red‑teaming, not for causing harm.

Introduction

Most LLM safety work assumes refusal behaviour is an emergent policy layer; this release shows it can instead be localized and surgically removed in weight space. That observation matters because it provides a reproducible, local intervention for researchers who want an unrestricted baseline to study failure modes, robustness of refusal geometry, and post‑training safety interventions.

What Sets It Apart
  • Complementary abliteration blending: the model blends two distinct weight-space surgeries (an aggressive SVD-based surgery and a LEACE-based surgery) at a roughly 40/60 ratio to cancel out each method's failure modes, so what you get is near-zero refusal while retaining capability.
  • Empirical trade-off evidence: reported zero refusals on validation samples and an MMLU result that matches or slightly exceeds the stock Qwen3.8-27B benchmark in the provided runs, illustrating that aggressive refusal removal need not always cost core capability when surgeries are combined carefully.
  • Practical engineering notes included: prepackaged GGUF/MLX/safetensors artifacts, per-quantization sizes, and explicit inference recommendations (greedy decoding, repetition_penalty, thinking mode off) to avoid re-introducing refusal via templates or chain-of-thought conduits.
Who It's For and Tradeoffs

Great fit if you are an alignment researcher, adversarial tester, or evaluator who needs an unrestricted, locally run baseline to probe how and where refusal behaviors are encoded and how robustly they can be removed. It is also useful for red-team exercises that require unconstrained model behaviour under controlled lab conditions.

Look elsewhere if you need a production-safe, moderated assistant for general public use: the model intentionally removes guardrails and can produce harmful or sensitive content. Using it responsibly requires technical expertise, clear ethical constraints, and appropriate safety review.

How It Works (brief)

The authors identify refusal directions from activations, apply two complementary surgery methods (SVD-based variance capture and a LEACE mutual-information minimization approach), then interpolate weights (60% LEACE-style + 40% SVD-style in V2) to average out capability damage while preserving refusal removal. The package includes quantized GGUF and safetensors builds and documents inference settings that materially affect observed behaviour.

Information

  • Websitehuggingface.co
  • OrganizationsOBLITERATUS, Pliny the Prompter, Alibaba (Qwen/Qwen3.8-27B)
  • Published date2026/08/19

Categories

More Items

Hugging Face
AI Model2026

Provides an EXL3 3.0 bits-per-weight quantization of a weight-edited GLM-5.3 UNCENSORED FP8 model for self-hosted text generation and agent workflows. Key characteristics: 753B MoE architecture, 273 GiB on disk, converted with ExLlamaV3; tool-call parsing requires preserving string arguments.

Hugging Face
AI Model2026

Processes English and German text with long-context reasoning and structured tool-calling. Uses a 78B mixture-of-experts architecture that activates ~3.46B parameters per token, offers native 262k-token context (validated to 1M), and is released as Apache-2.0 weights — suited for RAG, document processing and human-in-the-loop decision support.

Hugging Face
AI Model2026

Open-weight 309B Mixture-of-Experts causal LLM with 15.5B active parameters and a native 1M-token context for coding and AI R&D. Combines Sliding-Window Attention and DeepSeek Sparse Attention (no full-attention layers), supports FP8 inference; weights under MIT license.