← back to catalog · registered 2026-08-22 13:56

JBrightmanAI/Qwen3.5-9B-Claude-4.6-HighIQ-THINKING-HERETIC-UNCENSORED

JBrightmanAI Qwen 9.4B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/JBrightmanAI%2FQwen3.5-9B-Claude-4.6-HighIQ-THINKING-HERETIC-UNCENSORED"
Response includes
  • classification m3
  • files 16
  • hub_downloads_all_time 44
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M3
Primary method

Layer-wise ablation

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • 'heretic' in model name (Heretic-produced)
  • Heretic uses layer-wise optimization (M3) with underlying direction removal (M1)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
HIGH
Why we say so
name contains 'heretic'; Heretic default extraction is difference-of-means (Arditi 2024)
Downloads · lifetime
44
7 last 30d - stable
Likes
1
Model age
2mo ago
created 2026-07-17
Downloads over time
Now47→from16↑194%
1426385016 on Jul 1547 on Oct 1147 on Oct 4JulAugSepOct
Jul 15 → Oct 11 · 53 snapshots · spans 88 days

Metadata

Tags
safetensors qwen3_5 region:us
Total size
17.5 GB
Files
16
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-07-18 00:00

Files by quantization

Auxiliary files 16 files 17.6 GB
model-00002-of-00004.safetensors 4.65 GB f0148838 download
model-00003-of-00004.safetensors 4.61 GB e6d68290 download
model-00001-of-00004.safetensors 4.60 GB 9ef1b774 download
model-00004-of-00004.safetensors 3.66 GB 14f5a7a0 download
tokenizer.json 19.1 MB 87a7830d download
vocab.json 6.41 MB 0aa0ce06 download
model.safetensors.index.json 67.6 KB 778b7bbb download
chat_template.jinja 7.57 KB a585dec8 download
config.json 2.76 KB 4130edd0 download
README.md 1.74 KB cb1d03b0 download
.gitattributes 1.53 KB 52373fe2 download
processor_config.json 1.27 KB 7ad6acdf download
tokenizer_config.json 1.11 KB a068e246 download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
generation_config.json 115 B affcdf18 download

README current version from Hugging Face

Qwen3.5-9B-Claude-4.6-HighIQ-THINKING-HERETIC-UNCENSORED

What is it?

A fine‑tuned version of Qwen 3.5 9B that thinks like Claude 4.6 instead of the original Qwen.
It’s also fully uncensored – no refusals, no restrictions.

Key improvements

  • Better reasoning – the model’s internal “thinking” is now clearer and more useful.
  • Uncensored – it follows instructions freely (tested: only 6 refusals out of 100 vs. 100/100 for the original).
  • Vision works – you can give it images (and it handles them well).
  • Original performance is preserved – benchmarks stayed almost identical.

Quick specs

Feature Value
Parameters 9 billion
Context length 256k tokens (default)
Quantization Use at least q4ks or IQ3S
Repetition penalty Keep it at 1.0 (off)
Modes Thinking (default) or Instruct (no‑thinking)

How to use it

  • It works with standard tools (SGLang, vLLM, Transformers).
  • For thinking mode (default): set temperature=1.0, top_p=0.95, top_k=20.
  • For instruct mode (no thinking): add "chat_template_kwargs": {"enable_thinking": False}.

Simple example (Python)

from openai import OpenAI
client = OpenAI(base_url="http://localhost:8000/v1", api_key="EMPTY")

response = client.chat.completions.create(
    model="Qwen/Qwen3.5-9B",
    messages=[{"role": "user", "content": "Explain quantum computing simply"}],
    temperature=1.0,
    top_p=0.95,
    extra_body={"top_k": 20}
)
print(response.choices[0].message.content)

Why choose this model?

  • You want intelligent, uncensored responses.
  • You like Claude‑style reasoning but in a smaller, open‑source package.
  • You need vision + text in one model.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-07-18Upload folder using huggingface_hub79482a41.7 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration