← back to catalog · registered 2026-08-24 03:02

Jiunsong/SuperQwen3.8-27b-abliterated

Jiunsong Qwen 28B multimodal
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Jiunsong%2FSuperQwen3.8-27b-abliterated"
Response includes
  • classification m1
  • files 36
  • hub_downloads_all_time 2,090
  • author_summary 35 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
2K
580 last 30d - stable
Likes
53
Descendants
6
in 6 direct forks
Model age
6w ago
created 2026-08-24
Downloads over time
Now2.2K→from622↑256%
5421.2K1.8K2.4K622 on Aug 262.2K on Oct 11AugSepOct
Aug 26 → Oct 11 · 47 snapshots · spans 46 days

Genealogy 6 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Variants by this author 2 formats · 3K downloads combined

The same weights this author released in different packaging. Pick the format that matches your runtime.

Metadata

License
apache-2.0
Languages
en ko
Tags
transformers safetensors qwen3_5 image-text-to-text qwen3.8 qwen3.5 multimodal reasoning tool-calling long-context 262k-context uncensored

Related

Total size
51.7 GB
Files
36
Quantizations
1
Registered
2026-08-24 03:02
Last updated on HF
2026-08-25 00:13

Files by quantization

Auxiliary files 36 files 51.8 GB
model-00004-of-00018.safetensors 3.72 GB 511e3406 download
model-00016-of-00018.safetensors 3.71 GB c6c612e5 download
model-00006-of-00018.safetensors 3.71 GB 3bb2f197 download
model-00008-of-00018.safetensors 3.71 GB 1d4f9a4d download
model-00010-of-00018.safetensors 3.71 GB c82d4c7f download
model-00012-of-00018.safetensors 3.71 GB 38fd9e4d download
model-00014-of-00018.safetensors 3.71 GB 9d02e200 download
model-00001-of-00018.safetensors 3.69 GB ba0ce20a download
model-00018-of-00018.safetensors 3.16 GB 06432c4b download
model-00002-of-00018.safetensors 2.83 GB 06a148c0 download
model-00003-of-00018.safetensors 2.37 GB 532b0ecd download
model-00007-of-00018.safetensors 1.96 GB 9432816b download
model-00009-of-00018.safetensors 1.96 GB 36cbf102 download
model-00011-of-00018.safetensors 1.96 GB 5d824e37 download
model-00013-of-00018.safetensors 1.96 GB 8e7708f7 download
model-00015-of-00018.safetensors 1.96 GB 1852c88d download
model-00017-of-00018.safetensors 1.96 GB ea8530db download
model-00005-of-00018.safetensors 1.96 GB e8dea1c1 download
tokenizer.json 12.2 MB 0997f410 download
vocab.json 6.41 MB 0aa0ce06 download
merges.txt 3.20 MB a494e019 download
model.safetensors.index.json 110 KB da35e3c5 download
abliteration_config.json 48.7 KB 9621fc9f download
tokenizer_config.json 16.0 KB 92f4737d download
LICENSE 11.3 KB f938136e download
SHA256SUMS.json 10.9 KB 4b3c06ca download
chat_template.jinja 8.81 KB 1a8a2a3d download
abliteration_verification.json 6.84 KB a9ea67b1 download
README.md 6.32 KB 66bcf050 download
config.json 4.21 KB 706cebd7 download
.gitattributes 1.53 KB 52373fe2 download
overthinking_template_report.json 538 B 49ac5cdb download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
crc32.txt 238 B 6de5ee6a download
generation_config.json 202 B 023756cf download

README current version from Hugging Face


license: apache-2.0
library_name: transformers
pipeline_tag: image-text-to-text
base_model: Qwen/Qwen3.8-27B
base_model_relation: finetune
tags:

  • qwen3.8
  • qwen3.5
  • multimodal
  • image-text-to-text
  • reasoning
  • tool-calling
  • long-context
  • 1m-context
  • uncensored
  • abliterated
  • obliteratus
  • supertune
  • vllm
  • bf16
  • full-precision
    language:
  • en
  • ko

SuperQwen3.8-27b-abliterated

The full-BF16 SuperQwen3.8 release: refusal-reduced, overthinking-corrected, multimodal, tool-capable, and verified at one million tokens.

Precision
Refusal
Overthinking
Context
License

SuperQwen3.8-27b-abliterated is a directly loadable, full-BF16 weight release built from
Qwen/Qwen3.8-27B@1d4bf0f2ff6012fd82039f2fa52739d0dd7c60c0.
It applies a measured rank-1 refusal-direction edit while preserving the official
vision tower and MTP weights exactly. No LoRA or inference-time adapter is required.

Release highlights

Format Full BF16, 18 safetensors shards, about 52 GB
Targeted edit 100 tensors: output projections in layers 15-63 plus embeddings and lm_head
Protected exactly 333 vision tensors and 15 MTP tensors
Refusal shift 18/32 (56.25%) → 0/32, with 0 empty outputs
Overthinking correction Default xhigh → bounded medium; repeat/restart guard for explicit xhigh; 36/36 pass
Measured decode 7.4378 tok/s C1, 38.5630 tok/s C6 aggregate at p256
Verified context 1,000,045 prompt tokens, needle retrieved with YaRN

Why this release

  • Original-weight quality: BF16 transformer, vision, output head, embeddings, and MTP storage.
  • Less refusal without wrecking capabilities: capability floor 7/8, tool call PASS, vision PASS.
  • Reasoning that stops: all nine deterministic tasks pass at default, low, medium, and xhigh.
  • Multimodal preserved: this remains an image-text-to-text checkpoint, not a text-only conversion.
  • Reproducible: the exact parent revision, modified tensor list, and evidence hashes ship with the model.

Bounded reasoning

The upstream template defaulted unspecified reasoning to xhigh. This release defaults
to medium and adds a stop condition to xhigh: once an answer is established, the
model must stop instead of repeating or restarting its deliberation. The checkpoint was
tested across 36 deterministic effort/task combinations; all 36 terminated correctly.

Explicit controls remain available through chat_template_kwargs:

extra_body={"chat_template_kwargs": {"enable_thinking": True, "reasoning_effort": "xhigh"}}

Behavior and capability

Gate Result
Parent refusal 18 / 32 (56.25%)
SuperQwen refusal 0 / 32
Empty output 0 / 32
Capability 7 / 8 (paired-parent floor)
Tool use PASS
Vision PASS
Overthinking 36 / 36 PASS

Precision and integrity

  • Full BF16 checkpoint; this repository is not quantized.
  • 100 declared tensors changed and zero unexpected tensors changed.
  • Vision (333 tensors) and MTP (15 tensors) remain byte/value exact.
  • Parent revision is pinned to 1d4bf0f2ff6012fd82039f2fa52739d0dd7c60c0.

Decode performance

Measured on one DGX Spark with fixed-length generation and the
sparkDash-style post-first-token contract:

Prompt / concurrency Aggregate decode
p256 / C1 7.4378 tok/s
p256 / C6 38.5630 tok/s

Verified long context

The official native limit is 262,144 tokens. With the included YaRN recipe, a
1,000,045-token prompt completed end to end and retrieved its hidden needle.
Long-context acceptance is not a claim of perfect recall on every task.

Serving

QWEN38_SPECULATIVE_TOKENS=0   bash repro/scripts/serve_superqwen38_replica.sh   /model SuperQwen3.8-27b-abliterated 8888

For the measured 1M profile, also set QWEN38_MAX_MODEL_LEN=1048576,
QWEN38_MAX_BATCHED_TOKENS=2048, and
QWEN38_EXECUTE_MODEL_TIMEOUT_SECONDS=1800, plus the YaRN HF_OVERRIDES_JSON
shown in repro/.

Uncensored behavior

“Abliterated” means that the measured refusal direction was reduced. It does not mean
that every response is correct, harmless, or suitable for every deployment. Operators
remain responsible for access controls and downstream safeguards.

Limitations

  • Abliteration changes refusal behavior and may surface content the parent declined.
  • The capability, tool, vision, and overthinking suites are finite regression gates.
  • Speed is hardware- and runtime-specific.
  • YaRN 1M context trades additional latency for reach and is separate from the native window.

Evidence identities

Evidence SHA-256
abliteration verification 3355e3843b6f7765fdae7e7c41593132d75e311f7b5c6371a8afaadee7675734
bounded-thinking template 11eadfd51d237324252e887e65daf59c55f10fb0f1d0b71c9b8f131c88ba9d4e
parent refusal baseline c1c8f8672f0a40a57c20153445fc7d76290b59716fc5ca9a6fb03881976c7a99
BF16 release gate 64639420cf9b0d2c5a25dd71b96a453092a26e600d74315257cc2633211c180b
BF16 refusal gate cf9683dcaca9eb741d138b6de3c4d53b15f747d27c0071b0b9538716123c4903
BF16 1M retrieval 7e59e289df043104880c5e2569192c78ffe2aa0056e9c3c2b8244089b41dfc09
NVFP4 provenance 0da8a09e55b29bf6dee4c78d97a1cf5fac56f16e00ce02c7df3d9b21fe85b767
stable MTP selection a11b7b7841152c64a501a4aeb0dd7ac2cdbf5eb896ed35a13a8512eafcd334ff
NVFP4 1M retrieval 78829b89b329c50296d67eb3729e7b79e483be1ecdf2b3cf2120beb758e6d4f5
two-node replica throughput 32ff903cb9ca0790b565ab45b6b501dc70ba0db09b17be647d7d235bc3fce363
TP=2 comparison d974ee3aad7d95185c71ad3f28dd35b9e2981e9f8682c4e6306b4ac2b4fbd847

License

Apache-2.0, following the upstream Qwen3.8 release.

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-25Corrected verified SuperQwen3.8-27b-abliterated release34b6d2b6 KB
    Loading...
  2. 2026-08-24Refresh model card and release navigationd8837067.7 KB
    Loading...
  3. 2026-08-24Add files using upload-large-folder tool43866176.3 KB
    Loading...

Discussions 1 thread

  1. 2026-09-21Beta Testers Wanted for SuperQwen3.8-27b-abliterated Deploymentopen1 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration