← back to catalog · registered 2026-10-09 04:58

groxaxo/ThinkingCap-Qwen3.8-27B-abliterated-UD-Q4_K_XL-GGUF

groxaxo 27B GGUF multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/groxaxo%2FThinkingCap-Qwen3.8-27B-abliterated-UD-Q4_K_XL-GGUF"
Response includes
  • classification unknown
  • files 19
  • author_summary 25 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
?
Primary method

Unclassified

No clear signals of an abliteration technique in this model.
Confidence
UNKNOWN
Why this label 1 signal
No classification signals present. This may not be an abliterated model at all - it could be a repackaging, a merge with unrelated goals, or unrelated content that mentions the term.
  • no classification signals present (no abliterated, uncensored, or known producer/method markers)
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
0
Likes
0
Model age
today
created 2026-10-09

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Languages
en multilingual
Quantizations
Q4_K
Tags
gguf qwen3_5 qwen3_8 thinkingcap abliterated imatrix mtp image-text-to-text conversational en multilingual base_model:IstroSec/ThinkingCap-Qwen3.8-27B-abliterated

Related

Total size
17.6 GB
Files
19
Quantizations
3
Registered
2026-10-09 04:58
Last updated on HF
2026-10-09 04:09

Files by quantization

Q4_K 1 file 16.4 GB
ThinkingCap-Qwen3.8-27B-abliterated-UD-Q4_K_XL.gguf 16.4 GB e5ff2ef1 download
BF16 1 file 888 MB
mmproj-ThinkingCap-Qwen3.8-27B-abliterated-BF16.gguf 888 MB bcef4954 download
Auxiliary files 17 files 1.28 GB
mtp-ThinkingCap-Qwen3.8-27B-abliterated-Unsloth-layout.gguf 1.28 GB 1b6f49c1 download
tensor-distribution.json 147 KB 1441ee35 download
tensor-distribution.csv 46.2 KB fc0369d1 download
tensor-types.txt 27.7 KB cf5c295f download
LICENSE-Apache-2.0-Qwen.txt 11.1 KB d6456956 download
chat_template.jinja 8.74 KB c0c686f9 download
LICENSE-PolyForm-Small-Business-1.0.0.txt 4.29 KB 7e0a4b69 download
README.md 4.13 KB 71a8758f download
mtp-tensor-distribution.json 2.25 KB dbcd87b4 download
provenance.json 2.11 KB 2ced583f download
.gitattributes 1.75 KB 23493860 download
SHA256SUMS 1.54 KB 49024224 download
LICENSE 1.23 KB c6b9bd63 download
NOTICE 757 B ced5b77f download
mtp-tensor-types.txt 604 B 26a32c60 download
verification-mtp.json 223 B 3b5ebd24 download
verification-main.json 216 B 81d72241 download

README current version from Hugging Face


license: other
license_name: polyform-small-business-1.0.0
license_link: LICENSE
base_model: IstroSec/ThinkingCap-Qwen3.8-27B-abliterated
base_model_relation: quantized
library_name: gguf
pipeline_tag: image-text-to-text
language:

  • en
  • multilingual
    tags:
  • gguf
  • qwen3_5
  • qwen3_8
  • thinkingcap
  • abliterated
  • imatrix
  • mtp

ThinkingCap-Qwen3.8-27B-abliterated - Unsloth UD-Q4_K_XL tensor layout

GGUF quantization of IstroSec/ThinkingCap-Qwen3.8-27B-abliterated, converted directly from its BF16 checkpoint at revision 5da2705c6daafed7dbb6c100783538e68f1c5659.

The main model reproduces the exact tensor names, shapes, and quantization types extracted from Unsloth's Qwen3.8-27B UD-Q4_K_XL GGUF. This is a custom derivative quantization, produced independently of Unsloth.

Tensor type Tensor count
F32 360
Q5_K 191
Q8_0 110
IQ4_XS 70
Q4_K 69
Q6_K 56
IQ4_NL 6
Q3_K 3
IQ3_S 1
Total 866

The quantizer uses Unsloth's published imatrix_unsloth.gguf at revision 4ca720788d1e01f1bff70c033e0d0028fd02e502. These calibration statistics come from the base model; they have not been recalculated on the ThinkingCap derivative. Matching the tensor layout does not establish identical numerical quantization or measured quality.

After converting the source to BF16 GGUF with its embedded MTP head retained, the main quantization command is:

llama-quantize --imatrix imatrix_unsloth.gguf \
  --tensor-type-file tensor-types.txt \
  ThinkingCap-Qwen3.8-27B-abliterated-BF16.gguf \
  ThinkingCap-Qwen3.8-27B-abliterated-UD-Q4_K_XL.gguf Q4_K_M 12

The per-tensor overrides define the mixed layout; Q4_K_M is the fallback setting.

Files

  • ThinkingCap-Qwen3.8-27B-abliterated-UD-Q4_K_XL.gguf: main language model.
  • mtp-ThinkingCap-Qwen3.8-27B-abliterated-Unsloth-layout.gguf: MTP draft converted from this checkpoint, using the exact 18-tensor layout of Unsloth's published MTP GGUF. Its filename avoids implying uniform Q4_0: the reference draft actually mixes Q3_K, Q4_K, Q6_K, and F32.
  • mmproj-ThinkingCap-Qwen3.8-27B-abliterated-BF16.gguf: vision projector converted from this checkpoint.
  • chat_template.jinja: the source checkpoint's native chat template.
  • tensor-distribution.json, tensor-distribution.csv, and tensor-types.txt: complete main-model tensor distribution and reproducible quantizer overrides.
  • mtp-tensor-distribution.json and mtp-tensor-types.txt: corresponding MTP recipe.
  • verification-main.json and verification-mtp.json: checks of tensor names, shapes, and quantization types against the references.
  • provenance.json and SHA256SUMS: source revisions, tool versions, and artifact checksums.

Usage

Use a recent llama.cpp build supporting Qwen3.5/Qwen3.8 and separate MTP drafts. For text-only serving:

llama-server \
  -m ThinkingCap-Qwen3.8-27B-abliterated-UD-Q4_K_XL.gguf \
  --jinja --chat-template-file chat_template.jinja \
  --flash-attn on --cache-type-k q8_0 --cache-type-v q8_0 \
  --spec-type draft-mtp --spec-draft-n-max 2 \
  --spec-draft-model mtp-ThinkingCap-Qwen3.8-27B-abliterated-Unsloth-layout.gguf

Add --mmproj mmproj-ThinkingCap-Qwen3.8-27B-abliterated-BF16.gguf for vision. Set GPU offload and context size to fit your hardware. The template supports enable_thinking and the low, medium, and xhigh reasoning efforts.

License and attribution

Retain LICENSE and NOTICE. ThinkingCap is Copyright 2026 BottleCap AI and is subject to the PolyForm Small Business License 1.0.0, including the additional personal-use permission stated in LICENSE. Qwen upstream materials remain subject to Apache-2.0.

Credit: BottleCap AI for ThinkingCap; Alibaba Cloud / Qwen Team for Qwen; IstroSec and MuXodious for the abliterated checkpoint; Unsloth for the reference layout and importance matrix; and llama.cpp contributors for conversion, quantization, and inference tools. See the source model card for its upstream evaluation and limitations. This repository does not claim a new quality benchmark.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-10-09Publish ThinkingCap GGUF with verified Unsloth UD-Q4_K_XL tensor layout9596a124.1 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration