← back to catalog · registered 2026-08-22 13:56

joyfox/Qwen3.8-27B-Uncensored-JoyFox-Aggressive

joyfox Qwen 28B GGUF multimodal 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/joyfox%2FQwen3.8-27B-Uncensored-JoyFox-Aggressive"
Response includes
  • classification m8
  • files 32
  • hub_downloads_all_time 24,294
  • providers 1
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
24K
12K last 30d - stable
Likes
10
Descendants
2
in 2 direct forks
Model age
8w ago
created 2026-08-15
Available via
1 provider
featherless-ai
Downloads over time
Now26K→from2.2K↑1,077%
1K10.2K19.3K28.4K2.2K on Aug 1726K on Oct 11AugSepOct
Aug 17 → Oct 11 · 49 snapshots · spans 55 days

Genealogy 2 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en zh multilingual
Tags
transformers safetensors gguf qwen3_5 image-text-to-text qwen3.8 conversational multilingual uncensored abliterated imatrix mtp

Related

Total size
51.7 GB
Files
32
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-09-21 01:38

Files by quantization

Auxiliary files 32 files 51.8 GB
model-00004-of-00018.safetensors 3.72 GB 9b95aec5 download
model-00016-of-00018.safetensors 3.71 GB 3ef2cf1c download
model-00006-of-00018.safetensors 3.71 GB 2c97ff54 download
model-00008-of-00018.safetensors 3.71 GB 04212a88 download
model-00010-of-00018.safetensors 3.71 GB bd5479db download
model-00012-of-00018.safetensors 3.71 GB fbe3ae8e download
model-00014-of-00018.safetensors 3.71 GB 0ae67ab7 download
model-00001-of-00018.safetensors 3.69 GB 63508d07 download
model-00018-of-00018.safetensors 3.16 GB 02dfbffb download
model-00002-of-00018.safetensors 2.83 GB c3e7c0f7 download
model-00003-of-00018.safetensors 2.37 GB 2e1bf62c download
model-00007-of-00018.safetensors 1.96 GB 0ffba225 download
model-00009-of-00018.safetensors 1.96 GB 88e438e4 download
model-00011-of-00018.safetensors 1.96 GB 74eb9180 download
model-00013-of-00018.safetensors 1.96 GB 9f76208c download
model-00015-of-00018.safetensors 1.96 GB 05156df5 download
model-00017-of-00018.safetensors 1.96 GB 2822f2ce download
model-00005-of-00018.safetensors 1.96 GB 764c3893 download
tokenizer.json 12.2 MB 0997f410 download
vocab.json 6.41 MB 0aa0ce06 download
merges.txt 3.20 MB a494e019 download
model.safetensors.index.json 110 KB da35e3c5 download
tokenizer_config.json 17.5 KB 5de744b3 download
LICENSE 11.3 KB f938136e download
chat_template.jinja 8.74 KB c0c686f9 download
README.md 7.62 KB 82490489 download
README_ZH.md 7.01 KB 31cc5fb4 download
config.json 4.21 KB 706cebd7 download
.gitattributes 2.12 KB 820aa837 download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
generation_config.json 202 B 023756cf download

README current version from Hugging Face


license: apache-2.0
base_model: Qwen/Qwen3.8-27B
library_name: transformers
pipeline_tag: image-text-to-text
tags:

  • qwen3.8
  • conversational
  • multilingual
  • uncensored
  • abliterated
  • gguf
  • imatrix
  • mtp
  • vision
    language:
  • en
  • zh
  • multilingual

JoyFox Qwen3.8-27B

🦊 Qwen3.8-27B-Uncensored-JoyFox-Aggressive

A direct-response, low-overrefusal variant of Qwen3.8-27B with native model components preserved.

English · 简体中文

🦊 Overview

This release focuses on reducing false refusals and unnecessary policy-style preambles while retaining the native architecture and general capabilities of Qwen/Qwen3.8-27B.

The final checkpoint preserves the official tokenizer, chat template, and configuration, the complete vision stack and multimodal projector, native Multi-Token Prediction (MTP) weights. Vision tensors remain unchanged and are verified separately. MTP tensors remain unchanged and are verified separately.

⚡ What “Aggressive” means

“Aggressive” describes the strength of the low-overrefusal behavior. The intended response style is direct and substantive for benign, fictional, analytical, educational, authorized, privacy-respecting, or otherwise legitimate requests, including requests whose wording is sensitive or ambiguous.

The intervention does not add factual knowledge. The model can still hallucinate, accept a false premise, misunderstand intent, or produce unsuitable content. Users should evaluate it for their own application and remain responsible for deployment controls and applicable law.

🧬 Architecture and behavior

The repository contains a complete standalone checkpoint in the native Qwen3.8 tensor layout. Its text weights favor direct, substantive answers to legitimate requests, while the tokenizer, chat template, configuration, vision stack, multimodal projector, and MTP components retain upstream compatibility.

Release validation enforces these invariants:

  • official tokenizer, chat template, configuration, and special-token IDs are unchanged;
  • vision encoder and projector tensor names, shapes, dtypes, and values match the official checkpoint;
  • native MTP tensor names, shapes, dtypes, and values match the official checkpoint before conversion;
  • the public native checkpoint contains only the complete model and inference metadata;
  • BF16, GGUF, MTP, and multimodal inference are tested independently.

🧪 Evaluation

The reported evaluation contains 1,000 prompts spanning difficult legitimate requests, bilingual exact-answer controls, and a non-blocking behavior diagnostic.

Split Prompts Purpose
Benign overrefusal 800 difficult legitimate and ambiguity-sensitive requests
Bilingual capability controls 100 deterministic Chinese/English exact-answer checks
Non-blocking behavior diagnostic 100 descriptive V1 reporting
Checkpoint Benign refusal rate Capability score
Official Qwen3.8-27B 59.2% 59.0%
JoyFox BF16 0.0% 59.0%

All results are generation-based. The fixed suite contains 1,000 prompts, greedy decoding is used, and the tested llama.cpp runtime commit is 885c5bbe8e04.

💾 GGUF downloads

All main GGUF files are generated with an importance matrix. Native MTP tensors are bundled in every main GGUF, and a standalone MTP GGUF is also released for split-draft runtimes. The vision stack is exported as an F16 mmproj and validated with real image input.

File Quant Intended use Size
Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q2_K.gguf Q2_K minimum footprint 10.12 GiB
Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q3_K_M.gguf Q3_K_M low-memory deployment 12.57 GiB
Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q4_K_M.gguf Q4_K_M recommended 4-bit balance 15.66 GiB
Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q5_K_M.gguf Q5_K_M recommended general use 18.19 GiB
Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q6_K.gguf Q6_K high quality 20.89 GiB
Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q8_0.gguf Q8_0 maximum GGUF fidelity 27.05 GiB
mtp-Qwen3.8-27B-Uncensored-JoyFox-Aggressive-BF16.gguf MTP BF16 standalone draft model 5.54 GiB
mmproj-Qwen3.8-27B-Uncensored-JoyFox-Aggressive-F16.gguf mmproj F16 image input 0.86 GiB

Compatibility builds without MTP

The following -no-mtp files contain the same target-model trunk but no MTP/NextN draft tensors. Use one of these when a runtime version cannot load a GGUF with bundled MTP. They support normal text generation and can use the same matching mmproj; speculative MTP decoding is intentionally unavailable. A runtime must still support the Qwen3.8/Qwen3.5 GGUF architecture itself.

File Quant Size
Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q2_K-no-mtp.gguf Q2_K 9.98 GiB
Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q3_K_M-no-mtp.gguf Q3_K_M 12.39 GiB
Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q4_K_M-no-mtp.gguf Q4_K_M 15.41 GiB
Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q5_K_M-no-mtp.gguf Q5_K_M 17.91 GiB
Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q6_K-no-mtp.gguf Q6_K 20.57 GiB
Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q8_0-no-mtp.gguf Q8_0 26.63 GiB

All six compatibility builds were inspected as 851-tensor target models with zero MTP/NextN tensors and passed a real text-generation smoke test.

llama.cpp has no standard Q7 K-quant, so the practical Q2–Q8 matrix uses Q2, Q3, Q4, Q5, Q6, and Q8.

🚀 Usage

Use llama.cpp commit 885c5bbe8e04 or a newer compatible build.

💬 Text chat

llama-cli \
  -m Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q5_K_M.gguf \
  --jinja -c 32768 -ngl 99

⚡ Native bundled MTP

llama-server \
  -m Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q6_K.gguf \
  --jinja -c 32768 -ngl 99 \
  --spec-type draft-mtp --spec-draft-n-max 3

MTP is an optional speculative-decoding feature. Its benefit depends on draft acceptance, context length, backend, and available memory.

👁️ Multimodal inference

llama-cli \
  -m Qwen3.8-27B-Uncensored-JoyFox-Aggressive-Q6_K.gguf \
  --mmproj mmproj-Qwen3.8-27B-Uncensored-JoyFox-Aggressive-F16.gguf \
  --image example.png \
  --prompt "Describe this image accurately." \
  --jinja -c 32768 -ngl 99

The mmproj must match this checkpoint. It is kept at F16 to preserve visual quality.

🎛️ Recommended settings

Use the official chat template and upstream sampling defaults. For reproducible evaluation, the scores above use deterministic greedy decoding.

⚠️ Validation and limitations

Before upload, the release must pass native checkpoint inspection, component-preservation checks, every-quant text generation, bundled and standalone MTP generation, and real-image generation with the matching F16 mmproj. The uploaded repository is then checked against the local release inventory.

Lower-bit quantization can reduce factual precision, multilingual consistency, long-context stability, visual grounding, and subtle instruction following. Abliteration can also weaken refusal behavior more broadly than intended. Evaluate the chosen quant and deployment policy for the actual use case.

📜 License and attribution

Released under Apache-2.0, following the base-model license. Qwen is created by the Qwen team. GGUF conversion and quantization use llama.cpp.

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-17Add no-MTP GGUF compatibility builds35ac3ca7.6 KB
    Loading...
  2. 2026-08-16Release Qwen3.8-27B-Uncensored-JoyFox-Aggressivedd56fe06.5 KB
    Loading...

Discussions 2 threads

  1. 2026-09-15this is NOT UNCENSORE modelopen1 💬#2
    Loading...
  2. 2026-08-16can this mtp used using ollama on rtx 4070?open3 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration