← back to catalog · registered 2026-08-22 13:56

aday777/Qwen3.8-27B-ARA-abliterated-NVFP4-MTP

aday777 Qwen 24B multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/aday777%2FQwen3.8-27B-ARA-abliterated-NVFP4-MTP"
Response includes
  • classification m1
  • files 22
  • hub_downloads_all_time 13,793
  • author_summary 3 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
14K
Likes
2
Model age
8w ago
created 2026-08-15
Downloads over time
Now21.4K→from4.7K↑359%
3.8K10.2K16.6K23K4.7K on Aug 1921.4K on Oct 11AugSepOct
Aug 19 → Oct 11 · 48 snapshots · spans 53 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Quantizations
BF16
Tags
safetensors qwen3_5 qwen3.8 abliterated uncensored nvfp4 mtp multimodal compressed-tensors image-text-to-text conversational base_model:trohrbaugh/Qwen3.8-27B-heretic-ara

Related

Total size
19.1 GB
Files
22
Quantizations
2
Registered
2026-08-22 13:56
Last updated on HF
2026-08-27 22:33

Files by quantization

BF16 1 file 810 MB
model-mtp-bf16.safetensors 810 MB 90fa0e3e download
Auxiliary files 21 files 18.4 GB
model-00001-of-00004.safetensors 4.66 GB b031a957 download
model-00003-of-00004.safetensors 4.63 GB b5292fef download
model-00002-of-00004.safetensors 4.62 GB 51cbd444 download
model-00004-of-00004.safetensors 4.45 GB e9fbe4c3 download
tokenizer.json 19.1 MB 445a1c45 download
model.safetensors.index.json 272 KB 527ff211 download
calibration_snapshot.json 40.2 KB cb34358a download
build_qwen38_abliterated_nvfp4.py 23.0 KB 7a24775b download
config.json 15.3 KB 948256bd download
LICENSE 11.3 KB f938136e download
chat_template.jinja 8.74 KB c0c686f9 download
BUILD_MANIFEST.json 3.89 KB 7c2ed387 download
README.md 2.24 KB dfd83e73 download
VALIDATION_REPORT.json 2.12 KB 839763d9 download
SHA256SUMS 1.74 KB 0e58c2a4 download
.gitattributes 1.53 KB 52373fe2 download
vllm_uuid_sitecustomize.py 1.27 KB e00663b8 download
tokenizer_config.json 1.24 KB a9eacca6 download
processor_config.json 1.16 KB 33818c7f download
recipe.yaml 256 B 3bbc8398 download
generation_config.json 214 B 0bc3addd download

README current version from Hugging Face


license: apache-2.0
base_model: trohrbaugh/Qwen3.8-27B-heretic-ara
base_model_relation: quantized
pipeline_tag: image-text-to-text
tags:

  • qwen3_5
  • qwen3.8
  • abliterated
  • uncensored
  • nvfp4
  • mtp
  • multimodal
  • compressed-tensors

Qwen3.8-27B ARA Abliterated / Uncensored NVFP4 + MTP

This is a multimodal, refusal-ablated (commonly described as "uncensored") W4A4
NVFP4 derivative of the reproducible
trohrbaugh/Qwen3.8-27B-heretic-ara
BF16 checkpoint. It is intended for native Blackwell NVFP4 inference.

What is preserved

  • The vision tower, recurrent convolutions, language head, and all 15 native MTP
    tensors remain BF16.
  • The MTP tensors were grafted from the hash-verified source after Transformers
    serialization and verified bit-exact.
  • All 333 vision tensors were verified bit-exact against the BF16 source.
  • The language-model linear layers use compressed-tensors NVFP4 W4A4 group-16
    quantization.

Validation

This artifact passed its text-capability, image-vision, benign refusal-surface,
native MTP-acceptance, integrity, and clean-load gates on vLLM 0.23. See
BUILD_MANIFEST.json, VALIDATION_REPORT.json, and SHA256SUMS for exact
provenance and results. Video tensors/processors are preserved, but video input
was not part of the live runtime gate.

The live gate used Qwen3_5ForConditionalGeneration, native three-token MTP,
the FlashInfer CUTLASS NVFP4 kernel, an 8,192-token context, and an RTX PRO 6000
Blackwell GPU. Loaded model memory was approximately 19.53 GiB.

vLLM example

vllm serve aday777/Qwen3.8-27B-ARA-abliterated-NVFP4-MTP \
  --served-model-name Qwen3.8-27B-ARA-NVFP4-MTP \
  --max-model-len 8192 \
  --speculative-config '{"method":"mtp","num_speculative_tokens":3}'

Use a recent vLLM build with Qwen3.5 multimodal and compressed-tensors NVFP4
support. Native NVFP4 execution requires compatible Blackwell hardware and CUDA
runtime support.

Notes

"Abliterated" or "uncensored" describes the source checkpoint's refusal-ablation
process; it is not a guarantee that every prompt will receive a particular answer.
Users remain responsible for evaluating outputs and applying safeguards appropriate
to their deployment.

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-27Add Bitcoin donation addressdf03f972.4 KB
    Loading...
  2. 2026-08-15Add files using upload-large-folder tool16b43982.2 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration