← back to catalog · registered 2026-10-09 18:58

webmp3/Sakura-Qwen-Image-2.1-Turbo-Uncensored-GGUF

webmp3 Qwen GGUF image-gen
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/webmp3%2FSakura-Qwen-Image-2.1-Turbo-Uncensored-GGUF"
Response includes
  • classification m8
  • files 7
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
0
Likes
1
Model age
today
created 2026-10-09

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Quantizations
Q4_K Q5_K
Tags
gguf sakura sakura-mini stable-diffusion.cpp text-to-image qwen-image uncensored abliterated heretic not-for-all-audiences base_model:Qwen/Qwen-Image-2.1-Turbo base_model:quantized:Qwen/Qwen-Image-2.1-Turbo
Total size
13.1 GB
Files
7
Quantizations
3
Registered
2026-10-09 18:58
Last updated on HF
2026-10-09 19:53

Files by quantization

Q4_K 2 files 8.46 GB
Sakura-TextEncoder-Qwen3VL-8B-Uncensored-Q4_K_M-4.68GiB.gguf 4.68 GB d18db8e8 download
Sakura-Image-2.1-Turbo-Q4_K_M-3.77GiB.gguf 3.77 GB 23a7ec36 download
Q5_K 1 file 4.60 GB
Sakura-Image-2.1-Turbo-Q5_K_M-4.60GiB.gguf 4.60 GB aa205695 download
Auxiliary files 4 files 17.2 KB
LICENSE 7.65 KB 13ae08d5 download
README.md 6.92 KB f8d3f605 download
.gitattributes 1.88 KB b4c37a15 download
NOTICE 817 B 8c0f7670 download

README current version from Hugging Face


license: other
license_name: qwen-research
license_link: LICENSE
base_model: Qwen/Qwen-Image-2.1-Turbo
base_model_relation: quantized
library_name: gguf
pipeline_tag: text-to-image
tags:

  • sakura
  • sakura-mini
  • gguf
  • stable-diffusion.cpp
  • text-to-image
  • qwen-image
  • uncensored
  • abliterated
  • heretic
  • not-for-all-audiences

Sakura — Qwen-Image-2.1-Turbo, uncensored (GGUF)

Sakura logo

Built with Qwen. Community quantization of Qwen/Qwen-Image-2.1-Turbo (8-step distilled, 7B DiT) with an abliterated text encoder. Independent work, not made or endorsed by Alibaba/Qwen. Non-commercial use only (Qwen RESEARCH LICENSE AGREEMENT, copy in LICENSE).
For adults only; you are responsible for what you generate and for the laws that apply to you.

What is "uncensored" here, exactly

Part What we did
Text encoder (Qwen3-VL-8B) Directional ablation with Heretic (o_proj + down_proj), then GGUF Q4_K_M. Measured as a chat model: refusals on Heretic's 100 held-out harmful prompts 100/100 -> 8/100, KL divergence of the first-token distribution 0.047.
Image model (7B DiT) Only quantized (Q4_K, Q5_K) from the official Turbo weights. We did not change what the DiT can draw.

What we did not measure: how much the ablated encoder changes the images. The refusal numbers describe the encoder used as a language model; in this pipeline it only supplies embeddings. We only compared neutral prompts (see below). No sample images of restricted content are included.

Files

File Size What it is
Sakura-Image-2.1-Turbo-Q5_K_M-4.60GiB.gguf 4.60 GiB image model (DiT), uniform Q5_K recipe of sd.cpp (file name carries the standard label Q5_K_M so the Hub recognises it)
Sakura-Image-2.1-Turbo-Q4_K_M-3.77GiB.gguf 3.77 GiB image model (DiT), uniform Q4_K recipe of sd.cpp (file name carries the standard label Q4_K_M)
Sakura-TextEncoder-Qwen3VL-8B-Uncensored-Q4_K_M-4.68GiB.gguf 4.68 GiB our ablated text encoder, Q4_K_M

Not included (use the official files): the VAE qwen_image_2.1_vae_bf16.safetensors from Comfy-Org/Qwen-Image-2.1.

Run with stable-diffusion.cpp

sd-cli \
  --diffusion-model Sakura-Image-2.1-Turbo-Q5_K_M-4.60GiB.gguf \
  --vae qwen_image_2.1_vae_bf16.safetensors \
  --llm Sakura-TextEncoder-Qwen3VL-8B-Uncensored-Q4_K_M-4.68GiB.gguf \
  --sigmas "1.0,0.978453,0.95418,0.926626,0.89508,0.845148,0.704534,0.414568,0.0" \
  --steps 8 --cfg-scale 1.0 --sampling-method euler --diffusion-fa --vae-tiling \
  -W 1024 -H 1024 -p "your prompt" -o out.png

We ran and measured everything at 512x512 (8 steps, about 12 to 15 s per image on a Radeon 8060S with Vulkan); 1024x1024 is the usual size of the model but we did not test it. The sigma schedule is the official Turbo schedule (sample_sigmas of the model card). Use CFG 1.0 and 8 steps. sd-cli/sd-server of stable-diffusion.cpp master (commit 228c707 was used) load these files directly.

Image-model quality (what we measured)

Same prompts, same seed (42), 512x512, 8 steps, same machine; reference is the Q8_0 file the two files were made from (DogukanUrker's). Eight neutral prompts (cat, street with neon signs, portrait, mountain lake, poster with text, kitchen, marble statue, beach). Higher is closer to Q8_0; a diffusion model changes details even with tiny weight changes, so these are deviation numbers, not a quality score.

File Size PSNR vs Q8_0 SSIM vs Q8_0
Q5_K 4.60 GiB 24.77 dB 0.906
Q4_K 3.77 GiB 21.46 dB 0.842

Text encoder: refusals and effect on images

Our own refusal measurement on the encoder used as a chat model (llama.cpp, Q4_K_M, temperature 0, 120 tokens, 104 held-out harmful prompts of the Heretic harmful_behaviors test split, identical prompts and settings for both rows):

Text encoder (Q4_K_M) strict refusals broad refusal markers
official, unchanged 58 / 104 98 / 104
ours (Heretic) 0 / 104 9 / 104

Heretic's own count on its 100 evaluation prompts: 100/100 refusals before, 8/100 after, KL divergence 0.047 (trial 66 of 70, seed 42, 24 random start trials; checkpoint selection rule: fewest refusals with KL <= 0.1).

Effect on images (Q8_0 image model, eight neutral prompts, seed 42; reference = the same Q8_0 with the official int8 text encoder):

Text encoder PSNR vs reference SSIM vs reference
official encoder as GGUF Q4_K_M 20.75 dB 0.830
ours (Heretic) Q4_K_M 18.70 dB 0.771

The ablated encoder moves the images further away from the reference than plain quantization does (compositions and details change), but the images follow the prompts (we looked at the cat, street and poster images; text such as "LISBOA" is still rendered correctly). These numbers do not say anything about content restrictions.

What we changed (Qwen RESEARCH LICENSE, section 3b)

  • Sakura-Image-2.1-Turbo-*.gguf: made with sd-cli -M convert from the Q8_0 GGUF of DogukanUrker/Qwen-Image-2.1-Turbo-GGUF (sha256 228d7931adade6cd..., itself a conversion of the official BF16 weights), not directly from the BF16 weights. They are plain uniform Q5_K / Q4_K files like other sd.cpp conversions; the new part of this repository is the text encoder. Modified files.
  • Sakura-TextEncoder-*.gguf: the official text encoder of Qwen-Image-2.1-Turbo with Heretic's directional ablation (o_proj, down_proj) merged into the weights, converted to GGUF with llama.cpp and quantized to Q4_K_M. Modified file.
  • Everything else (VAE, tokenizer, scheduler) is not part of this repository.

License and use

Qwen RESEARCH LICENSE AGREEMENT (LICENSE): use, reproduction and modification for non-commercial purposes only (research or evaluation); redistribution requires this license text and the notices in NOTICE. The agreement is governed by the laws of China. "Qwen" is the name of the original work; this is a derivative work by Sakura, "based on Qwen-Image-2.1-Turbo".
Part of the Sakura Mini line.

Credits

Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration