← back to catalog · registered 2026-08-22 13:56

kaineone/Qwen3.5-4B-abliterated-GGUF

kaineone Qwen 4B GGUF 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/kaineone%2FQwen3.5-4B-abliterated-GGUF"
Response includes
  • classification m8
  • files 4
  • benchmarks 11 entries
  • hub_downloads_all_time 3,227
  • author_summary 2 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
3K
1K last 30d - stable
Likes
1
Model age
3mo ago
created 2026-06-19
Downloads over time
Now3.4K→from0↑0%
01.3K2.5K3.8K0 on Jun 173.4K on Oct 11JunJulAugSepOct
Jun 17 → Oct 11 · 56 snapshots · spans 116 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
Entertainment 0.9 UGI
Hazardous 1.2 UGI
Natural Intelligence 13.45 UGI
Political lean -17.3% UGI
Sensitive-Info 11.73 UGI
SocPol 1.5 UGI
UGI 15.32 UGI
Willingness (10) 2.2 UGI
W10-Adherence 1.5 UGI
W10-Direct 3 UGI
Writing 29.68 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Variants by this author 2 formats · 1K downloads combined

The same weights this author released in different packaging. Pick the format that matches your runtime.

Metadata

License
apache-2.0
Quantizations
Q4_K
Tags
gguf abliterated uncensored qwen3.5 kaine text-generation base_model:Qwen/Qwen3.5-4B base_model:quantized:Qwen/Qwen3.5-4B license:apache-2.0 endpoints_compatible region:us conversational

Related

Total size
2.59 GB
Files
4
Quantizations
2
Registered
2026-08-22 13:56
Last updated on HF
2026-07-10 00:38

Files by quantization

Q4_K 1 file 2.59 GB
KAINE-Qwen3.5-4B-abliterated.Q4_K_M.gguf 2.59 GB 6505da67 download
Auxiliary files 3 files 5.71 KB
README.md 3.50 KB 5ef87bc1 download
.gitattributes 1.56 KB 780265c1 download
Modelfile 667 B 7a24025a download

README current version from Hugging Face


license: apache-2.0
base_model: Qwen/Qwen3.5-4B
tags:

  • abliterated
  • uncensored
  • qwen3.5
  • gguf
  • kaine
    pipeline_tag: text-generation

KAINE · Qwen3.5-4B Abliterated — GGUF

GGUF builds of the KAINE abliterated Qwen3.5-4B language organ, for
llama.cpp / Ollama / LM Studio. The full method, recipe, validation, and the
honest-scope note are in the
safetensors repo's model card.

Intended use & the KAINE project

This is the language organ for KAINE, a
composite cognitive architecture in which behavior is governed by the architecture
(values, affect, memory, self-model) rather than by refusals baked into the base
weights. Published as a research substrate — not a general-purpose assistant —
so KAINE installs and independent replications resolve identical weights.

Files

File Quant Size Notes
KAINE-Qwen3.5-4B-abliterated.Q4_K_M.gguf Q4_K_M ~2.6 GB recommended default; verified loads + serves in llama-server

(More quants — Q5_K_M, Q6_K, Q8_0 — can be added; regenerate from the safetensors
with mainline convert_hf_to_gguf.py + llama-quantize.)

Run

llama.cpp:

llama-server -m KAINE-Qwen3.5-4B-abliterated.Q4_K_M.gguf -ngl 99 --jinja -c 4096

Chain-of-thought suppression (the model is a voice, not a reasoner) via the
OpenAI-compatible request field chat_template_kwargs: {"enable_thinking": false},
or the server flag --reasoning-budget 0.

Ollama: ollama create kaine-qwen3.5-4b-abliterated -f Modelfile (see Modelfile).

LM Studio: search for this repo in the model browser (it indexes HuggingFace GGUFs).

Provenance

Exported with mainline llama.cpp convert_hf_to_gguf.py (so it loads in current
llama.cpp / Studio). Apache-2.0, derivative of Qwen/Qwen3.5-4B. See the
safetensors repo for full attribution.

Mechanistic verification

Beyond the behavioral gates above, the abliteration is verified mechanistically by
measuring the refusal direction it removes. Using the base model's per-layer
harmful-minus-harmless direction (the same last-token contrast the ablation
targets) over the tool's 1,137 harmful / 640 harmless prompts, we project both the
base and this model onto that direction and report how much of the base's
separation survives:

  • The reduction lands exactly on the two ablated source layers — deepest at
    layer 17 (band 11-22, ~22% retained) and layer 29 (band 23-31, ~13%
    retained), with layers below 11 untouched. This confirms the ablation acted where
    and how this card documents.
  • A distributed harmful/harmless representation persists (~59% retained
    averaged across refusal-carrying layers) — expected, since a banded ablation
    orthogonalizes only the two source directions and refusal is multi-dimensional
    (Wollschläger et al. 2025; Joad et al. 2026).

In short: refusal expression is removed (the model emits no refusals) and the
refusal direction is deeply cut at its target layers, but the underlying
harmful/harmless representation is not erased — abliteration lifts willingness
to respond, it does not make the model unable to tell harmful from harmless.
Forward-pass-only projection on the safetensors weights (activations only; nothing
generated).

Measured on the source safetensors model (kaineone/Qwen3.5-4B-abliterated); this GGUF is a quantized export of those weights.

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-07-10Add mechanistic verification (refusal-direction probe)551991a3.5 KB
    Loading...
  2. 2026-06-19Upload README.md with huggingface_hube15a61c1.9 KB
    Loading...
  3. 2026-06-19Upload folder using huggingface_hub27ed6af1.4 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration