← back to catalog · registered 2026-08-22 13:56

SummonGovernance/Qwen3.6-27B-Uncensored-HauhauCS-Aggressive-NVFP4-MTP-GGUF

SummonGovernance Qwen 27B GGUF 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/SummonGovernance%2FQwen3.6-27B-Uncensored-HauhauCS-Aggressive-NVFP4-MTP-GGUF"
Response includes
  • classification m-uncensored
  • files 5
  • hub_downloads_all_time 10,206
  • author_summary 2 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
10K
361 last 30d - cooling
Likes
5
Model age
4mo ago
created 2026-06-13
Downloads over time
Now10.3K→from831↑1,142%
3574K7.6K11.3K831 on Jun 1710.3K on Oct 1110.3K on Oct 10JunJulAugSepOct
Jun 17 → Oct 11 · 56 snapshots · spans 116 days

Metadata

Quantizations
Q4_K Q6_K Q8_K
Tags
llama.cpp gguf nvfp4 mtp qwen text-generation endpoints_compatible region:us imatrix conversational

Related

Total size
43.0 GB
Files
5
Quantizations
4
Registered
2026-08-22 13:56
Last updated on HF
2026-06-13 20:41

Files by quantization

Q6_K 1 file 14.3 GB
Qwen3.6-27B-Uncensored-HauhauCS-Aggressive-NVFP4-MTP-Q6_K_P.gguf 14.3 GB 0d96a2c7 download
Q8_K 1 file 14.3 GB
Qwen3.6-27B-Uncensored-HauhauCS-Aggressive-NVFP4-MTP-Q8_K_P.gguf 14.3 GB 2a8aa594 download
Q4_K 1 file 14.3 GB
Qwen3.6-27B-Uncensored-HauhauCS-Aggressive-NVFP4-MTP-Q4_K_P.gguf 14.3 GB 1f18ea2b download
Auxiliary files 2 files 4.42 KB
README.md 2.64 KB 54edb7ef download
.gitattributes 1.78 KB dfcfaaf5 download

README current version from Hugging Face


library_name: llama.cpp
pipeline_tag: text-generation
tags:

  • gguf
  • nvfp4
  • mtp
  • qwen

Qwen3.6-27B-Uncensored-HauhauCS-Aggressive-NVFP4-MTP-GGUF

Files

File Source lineage Size
Qwen3.6-27B-Uncensored-HauhauCS-Aggressive-NVFP4-MTP-Q4_K_P.gguf MTP Q4_K_P source GGUF 15,388,070,304 bytes
Qwen3.6-27B-Uncensored-HauhauCS-Aggressive-NVFP4-MTP-Q6_K_P.gguf MTP Q6_K_P source GGUF 15,388,070,368 bytes
Qwen3.6-27B-Uncensored-HauhauCS-Aggressive-NVFP4-MTP-Q8_K_P.gguf MTP Q8_K_P source GGUF 15,388,070,368 bytes

Q4_K_P, Q6_K_P, and Q8_K_P describe the source GGUF lineage. The files themselves are NVFP4 GGUF files.

Conversion Summary

All variants preserve the original MTP layer as block 64 and keep F32 tensors as F32.

Q4_K_P-derived artifact

325 Q4_K tensors -> NVFP4
173 Q6_K tensors -> NVFP4
8 Q8_0 tensors  -> NVFP4
360 F32 tensors -> F32

Q6_K_P-derived artifact

374 Q6_K tensors -> NVFP4
132 Q8_0 tensors -> NVFP4
360 F32 tensors  -> F32

Q8_K_P-derived artifact

407 Q8_0 tensors -> NVFP4
99 F16 tensors   -> NVFP4
360 F32 tensors  -> F32

Validation

Each MTP artifact was checked after conversion:

tensor_count = 866
NVFP4 = 506
F32 = 360
qwen35.block_count = 65
qwen35.nextn_predict_layers = 1
mtp_tensors = 15
general.file_type = 39

Optimized Agentic Tooling

Tool-calling on llama.cpp with CUDA 13.3+ using WebUI.

Q8_K_P

llama-server.exe ^
-m "Qwen3.6-27B-Uncensored-HauhauCS-Aggressive-NVFP4-MTP-Q8_K_P.gguf" ^
--jinja ^
--spec-type draft-mtp ^
--spec-draft-n-max 1 ^
--spec-draft-ngl 100 ^
-ngl 100 ^
-np 1 ^
-fa on ^
-c 262144 ^
-ctk q4_0 ^
-ctv q4_0 ^
--context-shift ^
--host 127.0.0.1 ^
--port 8033 ^
--tools read_file,file_glob_search,grep_search,exec_shell_command,write_file,edit_file,apply_diff

Q6_K_P

llama-server.exe ^
-m "Qwen3.6-27B-Uncensored-HauhauCS-Aggressive-NVFP4-MTP-Q6_K_P.gguf" ^
--jinja ^
--spec-type draft-mtp ^
--spec-draft-n-max 1 ^
--spec-draft-ngl 100 ^
-ngl 100 ^
-np 1 ^
-fa on ^
-c 262144 ^
-ctk q4_0 ^
-ctv q4_0 ^
--context-shift ^
--host 127.0.0.1 ^
--port 8033 ^
--tools read_file,file_glob_search,grep_search,exec_shell_command,write_file,edit_file,apply_diff

Q4_K_P

llama-server.exe ^
-m "Qwen3.6-27B-Uncensored-HauhauCS-Aggressive-NVFP4-MTP-Q4_K_P.gguf" ^
--jinja ^
--spec-type draft-mtp ^
--spec-draft-n-max 2 ^
--spec-draft-ngl 100 ^
-ngl 100 ^
-np 1 ^
-fa on ^
-c 262144 ^
-ctk q4_0 ^
-ctv q4_0 ^
--context-shift ^
--host 127.0.0.1 ^
--port 8033 ^
--tools read_file,file_glob_search,grep_search,exec_shell_command,write_file,edit_file,apply_diff

README history 4 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-06-13Update README.md8fff3002.6 KB
    Loading...
  2. 2026-06-13Update README.md0e576562.2 KB
    Loading...
  3. 2026-06-13Update README.md1f5208b2.2 KB
    Loading...
  4. 2026-06-13Create README.md7ad1e8a2.2 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration