← back to catalog · registered 2026-09-14 02:56

Honkware/Qwen3.8-27B-EfficientThink-Uncensored-K3-Opus5-Grok4.6-GPT5.6Sol-SFT-SimPO-DFlash2-exl3-4.0bpw

Honkware 6.7B second-order
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
169
Likes
3
Model age
2d ago
created 2026-09-14
Downloads over time
Now169from0↑0%
0621241860 on Sep 14169 on Sep 16Sep
Sep 14 → Sep 16 · 3 snapshots · spans 2 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Tags
exllamav3 safetensors qwen3_5 exl3 quantized text-generation conversational base_model:nerkyor/Qwen3.8-27B-EfficientThink-Uncensored-K3-Opus5-Grok4.6-GPT5.6Sol-SFT-SimPO-DFlash2 base_model:quantized:nerkyor/Qwen3.8-27B-EfficientThink-Uncensored-K3-Opus5-Grok4.6-GPT5.6Sol-SFT-SimPO-DFlash2 license:apache-2.0 4-bit region:us

Related

Total size
15.2 GB
Files
17
Quantizations
1
Registered
2026-09-14 02:56
Last updated on HF
2026-09-14 02:31

Files by quantization

Auxiliary files 17 files 15.2 GB
model-00001-of-00002.safetensors 7.89 GB 70f9d94f download
model-00002-of-00002.safetensors 7.29 GB cb6d56eb download
tokenizer.json 12.2 MB 0997f410 download
vocab.json 6.41 MB 0aa0ce06 download
merges.txt 3.20 MB a494e019 download
quantization_config.json 614 KB aee318fc download
model.safetensors.index.json 285 KB 0f84112e download
tokenizer_config.json 17.5 KB 5de744b3 download
chat_template.jinja 8.74 KB c0c686f9 download
README.md 4.54 KB 6bf2a527 download
manifest.json 4.09 KB d1e28fc1 download
config.json 3.88 KB 4a0f80e6 download
SHA256SUMS 2.00 KB 4a74daa6 download
.gitattributes 1.53 KB 52373fe2 download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
generation_config.json 214 B 8b9f95da download

README current version from Hugging Face


license: apache-2.0
base_model: nerkyor/Qwen3.8-27B-EfficientThink-Uncensored-K3-Opus5-Grok4.6-GPT5.6Sol-SFT-SimPO-DFlash2
base_model_relation: quantized
quantized_by: Honkware
library_name: exllamav3
pipeline_tag: text-generation
tags:

  • exl3
  • exllamav3
  • quantized

quantization_format: exl3
inference: false
bits_per_weight: 4.0

Qwen3.8 · 27B · EfficientThink · Uncensored · K3 · Opus5 · Grok4.6 · GPT5.6Sol · SFT · SimPO · DFlash2

EXL3  ·  4.0 bpw  ·  16.3 GB  ·  Dense


format
bpw
size
codebook
arch

base model
quantized by
collection


[!NOTE]
An ExLlamaV3 build of nerkyor/Qwen3.8-27B-EfficientThink-Uncensored-K3-Opus5-Grok4.6-GPT5.6Sol-SFT-SimPO-DFlash2 at 4.0 bits per weight. See Quants for sibling repos at other bit‑widths or browse the collection.

Quants

BPW     Size     Status
4.0 16.3 GB this repo

Inference

Loader Use it for
TabbyAPI OpenAI‑compatible HTTP server. Drop‑in for OpenAI clients.
text‑generation‑webui Local chat UI. Pick the ExLlamaV3 loader from the model dropdown.
ExLlamaV3 Direct Python API for embedding the model in your own code or pipeline.

Download

pip install -U huggingface_hub

hf download \
  Honkware/Qwen3.8-27B-EfficientThink-Uncensored-K3-Opus5-Grok4.6-GPT5.6Sol-SFT-SimPO-DFlash2-exl3-4.0bpw \
  --local-dir ./Qwen3.8-27B-EfficientThink-Uncensored-K3-Opus5-Grok4.6-GPT5.6Sol-SFT-SimPO-DFlash2-exl3-4.0bpw
Quantization recipe  (advanced, embedded in quantization_config.json)
Setting Value
Format EXL3
Bits per weight 4.0
Head bits 6
Calibration rows 250
Calibration data exllamav3 bundled mix (c4, code, multilingual, technical, tiny, wiki)
Codebook mul1
Out‑scales always
Parallel mode enabled

License & use

[!IMPORTANT]
Use and license follow the base model.
Quantization adds no additional restrictions. Refer to the upstream repository for terms, citation, and safety documentation.


Quantized with BlockQuant  ·  convention {org}/{model}-exl3-{bpw}bpw

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-09-14docs: proper model card with base-model link + inference notes164ec304.5 KB
    Loading...
  2. 2026-09-14Upload README.md with huggingface_hub44cb4f04.6 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in app" button that hands off directly to a local runtime of your choice - Infrahuman, LM Studio, or Ollama. No API keys, no subscription, no prompt leakage.