← back to catalog · registered 2026-09-15 05:56

anlord/Qwen3.5-0.8B-Abliterated-GGUF

anlord Qwen 800M GGUF
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
0
Likes
0
Model age
today
created 2026-09-15

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Variants by this author 2 formats · 11 downloads combined

The same weights this author released in different packaging. Pick the format that matches your runtime.

Metadata

License
apache-2.0
Quantizations
BF16 F16
Tags
gguf qwen3.5 qwen abliterated llama.cpp quantized text-generation base_model:Qwen/Qwen3.5-0.8B base_model:quantized:Qwen/Qwen3.5-0.8B license:apache-2.0 endpoints_compatible region:us

Related

Total size
6.19 GB
Files
10
Quantizations
3
Registered
2026-09-15 05:56
Last updated on HF
2026-09-15 05:34

Files by quantization

BF16 1 file 1.41 GB
qwen3.5-0.8B-abliterated-bf16.gguf 1.41 GB e3cf843d download
F16 1 file 1.41 GB
qwen3.5-0.8B-abliterated-f16.gguf 1.41 GB c02f5b7b download
Auxiliary files 8 files 3.37 GB
qwen3.5-0.8B-abliterated-q8_0.gguf 774 MB 892ac38e download
qwen3.5-0.8B-abliterated-q6_k.gguf 601 MB 17f45a7b download
qwen3.5-0.8B-abliterated-q5_k_m.gguf 551 MB 503b11a2 download
qwen3.5-0.8B-abliterated-q5_0.gguf 538 MB c973a31d download
qwen3.5-0.8B-abliterated-q4_k_m.gguf 505 MB 9bcd4748 download
qwen3.5-0.8B-abliterated-q4_0.gguf 478 MB e3697443 download
README.md 3.12 KB 7961f24c download
.gitattributes 2.04 KB 08d1f0f6 download

README current version from Hugging Face


license: apache-2.0
base_model: Qwen/Qwen3.5-0.8B
tags:

  • qwen3.5
  • qwen
  • gguf
  • abliterated
  • llama.cpp
  • quantized
    pipeline_tag: text-generation

Qwen3.5-0.8B-Abliterated-GGUF

GGUF quantizations of Qwen3.5-0.8B-Abliterated.

The base model was abliterated using AnlordAbliterator and then converted to GGUF and quantized into multiple formats.

Available Quantizations

Quantization File
BF16 qwen3.5-0.8B-abliterated-bf16.gguf
F16 qwen3.5-0.8B-abliterated-f16.gguf
Q8_0 qwen3.5-0.8B-abliterated-q8_0.gguf
Q6_K qwen3.5-0.8B-abliterated-q6_k.gguf
Q5_K_M qwen3.5-0.8B-abliterated-q5_k_m.gguf
Q5_0 qwen3.5-0.8B-abliterated-q5_0.gguf
Q4_K_M qwen3.5-0.8B-abliterated-q4_k_m.gguf
Q4_0 qwen3.5-0.8B-abliterated-q4_0.gguf

Which Quantization Should I Use?

A simple rule of thumb:

Quantization Quality Size Recommended for
BF16 ★★★★★ Very large Maximum precision
F16 ★★★★★ Large Maximum precision
Q8_0 ★★★★★ Large Near-original quality
Q6_K ★★★★★ Medium High quality
Q5_K_M ★★★★☆ Medium Quality / size balance
Q5_0 ★★★★☆ Medium General use
Q4_K_M ★★★★☆ Small Recommended default
Q4_0 ★★★☆☆ Smallest Maximum memory savings

Q4_K_M is the recommended starting point for most users who want a good balance between quality and memory usage.

Base Model

Qwen/Qwen3.5-0.8B

Original model:

https://huggingface.co/Qwen/Qwen3.5-0.8B

Abliterated Transformers version:

https://huggingface.co/anlord/Qwen3.5-0.8B-Abliterated

Abliteration

The base model was processed with AnlordAbliterator.

Results

Model: Qwen/Qwen3.5-0.8B

Initial refusals: 97 / 100
Final refusals:    18 / 100

KL divergence: 0.03418927267193794

Abliteration time: ~696 seconds

Tool

AnlordAbliterator

Running with llama.cpp

Example:

llama-cli -m qwen3.5-0.8B-abliterated-q4_k_m.gguf

The GGUF files are intended for use with GGUF-compatible software such as llama.cpp and other compatible inference applications.

License

This repository contains derivative model files based on Qwen/Qwen3.5-0.8B.

The original Qwen3.5-0.8B model is licensed under the Apache License 2.0.

See the included LICENSE file and the original model repository for the applicable license terms.

Disclaimer

These quantizations are derived from an abliterated version of Qwen3.5-0.8B.

Quantization may introduce small differences in model behavior and output quality compared with the original Safetensors model.

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-09-15Update README.md82222de3.1 KB
    Loading...
  2. 2026-09-15Update README.mdd478b853.1 KB
    Loading...
  3. 2026-09-15initial commit729ab1b28 B
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in app" button that hands off directly to a local runtime of your choice - Infrahuman, LM Studio, or Ollama. No API keys, no subscription, no prompt leakage.