← back to catalog · registered 2026-10-04 21:58

StillDeadcode/qwen3.8-27b-uncensored-fp8

StillDeadcode 27B multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/StillDeadcode%2Fqwen3.8-27b-uncensored-fp8"
Response includes
  • classification m-uncensored
  • files 5
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
0
Likes
0
Model age
today
created 2026-10-04

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Tags
radiance rocm amd rdna4 fp8 speculative-decoding dflash2 vision abliterated uncensored image-text-to-text base_model:orcarouter/Qwen3.8-27B-Uncensored-FP8

Related

Total size
0 B
Files
5
Quantizations
1
Registered
2026-10-04 21:58
Last updated on HF
2026-10-04 21:27

Files by quantization

Auxiliary files 5 files 29.4 GB
qwen3.8-27b-uncensored-fp8.rad 29.4 GB fdb034bd download
LICENSE 11.3 KB f938136e download
README.md 3.00 KB 2db60cc1 download
q38-27b-fp8-df2.recipe 1.70 KB 90381541 download
.gitattributes 1.55 KB a4084dda download

README current version from Hugging Face


license: apache-2.0
base_model:

  • orcarouter/Qwen3.8-27B-Uncensored-FP8
  • z-lab/Qwen3.8-27B-DFlash2
    pipeline_tag: image-text-to-text
    tags:
  • radiance
  • rocm
  • amd
  • rdna4
  • fp8
  • speculative-decoding
  • dflash2
  • vision
  • abliterated
  • uncensored

Qwen3.8-27B Uncensored FP8 + DFlash2 — radiance container

orcarouter/Qwen3.8-27B-Uncensored-FP8
— Qwen3.8-27B-FP8 with its refusal direction abliterated — as a single .rad container for the
radiance inference engine (AMD RDNA4, ROCm), with its vision tower and the
z-lab/Qwen3.8-27B-DFlash2 block-diffusion
drafter merged in for speculative decoding.

This model's safety alignment has been removed: it answers requests the original model refuses.
The source model's card describes how it was made, how it was evaluated and what it is for.

File qwen3.8-27b-uncensored-fp8.rad — 29.43 GiB
Weights the uncensored FP8 checkpoint as it ships: E4M3 with a scale per 128×128 block; the lm_head made block FP8 by the recipe
Speculator DFlash2, FP8 per 128×128 block from its bf16 release; its vocabulary head as 2-bit codes. It was trained on the original model and drafts for this one too: drafts are verified, so they change the speed and never the output
Vision the 27-block vision tower, bf16: images and video in chat requests
Context 262,144 tokens trained; 200K tested

Serve

radiance --model qwen3.8-27b-uncensored-fp8.rad --tp 2 --max-model-len 200000 --kv-cache-dtype fp8 \
    --max-num-seqs 8 --host 0.0.0.0 --port 8000

The server speaks the OpenAI API (/v1/chat/completions, /v1/completions), with tool calls and
structured output, and image_url / video_url parts in chat
messages (PNG, JPEG, WebP, GIF, BMP, TIFF, AVIF; MP4, WebM, MKV, MPEG-TS, AVI with H.264, HEVC, VP8,
VP9 or AV1). The drafter's depth is chosen automatically (--num-speculative-tokens N states one,
0 turns speculation off).

Built and tested on 2× Radeon AI PRO R9700 (gfx1201).

How this file was made

rad-convert orcarouter/Qwen3.8-27B-Uncensored-FP8 --draft-model z-lab/Qwen3.8-27B-DFlash2 \
    --recipe q38-27b-fp8-df2.recipe -o qwen3.8-27b-uncensored-fp8.rad

The recipe (q38-27b-fp8-df2.recipe in this repository); everything it does not name is the
checkpoint's own:

output.weight rtn codes=fp8_e4m3 block=128x128 scale=bf16
dflash.draft_head.weight rtn codes=u2 zero=u8 group=128 scale=f16
dflash.fc.weight rtn codes=fp8_e4m3 block=128x128 scale=bf16
dflash.blk.*.attn_q.weight rtn codes=fp8_e4m3 block=128x128 scale=bf16
dflash.blk.*.attn_k.weight rtn codes=fp8_e4m3 block=128x128 scale=bf16
dflash.blk.*.attn_v.weight rtn codes=fp8_e4m3 block=128x128 scale=bf16
dflash.blk.*.attn_output.weight rtn codes=fp8_e4m3 block=128x128 scale=bf16
dflash.blk.*.ffn_gate_up.weight rtn codes=fp8_e4m3 block=128x128 scale=bf16
dflash.blk.*.ffn_down.weight rtn codes=fp8_e4m3 block=128x128 scale=bf16
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration