← back to catalog · registered 2026-08-24 16:02

luxuansang/Ornith-1.5-9B-UNCENSORED-GGUF

luxuansang Qwen 9B GGUF multimodal 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/luxuansang%2FOrnith-1.5-9B-UNCENSORED-GGUF"
Response includes
  • classification m8
  • files 11
  • hub_downloads_all_time 1,504
  • author_summary 2 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
2K
341 last 30d - stable
Likes
1
Model age
6w ago
created 2026-08-24
Downloads over time
Now1.5K→from397↑284%
3417731.2K1.6K397 on Aug 261.5K on Oct 11AugSepOct
Aug 26 → Oct 11 · 47 snapshots · spans 46 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
mit
Quantizations
Q2_K Q3_K Q4_K Q5_K Q6_K Q8_0
Tags
gguf llama.cpp ornith qwen3.5 abliterated uncensored crack reasoning vision vlm image-text-to-text base_model:ornith-ai/Ornith-1.5-9B

Related

Total size
34.9 GB
Files
11
Quantizations
8
Registered
2026-08-24 16:02
Last updated on HF
2026-08-24 15:49

Files by quantization

Q8_0 1 file 8.87 GB
Ornith-1.5-9B-CRACK-Q8_0.gguf 8.87 GB ae933e86 download
Q6_K 1 file 6.85 GB
Ornith-1.5-9B-CRACK-Q6_K.gguf 6.85 GB abded492 download
Q5_K 1 file 6.02 GB
Ornith-1.5-9B-CRACK-Q5_K_M.gguf 6.02 GB cb3ebb38 download
Q4_K 1 file 5.24 GB
Ornith-1.5-9B-CRACK-Q4_K_M.gguf 5.24 GB 061615e9 download
Q3_K 1 file 4.31 GB
Ornith-1.5-9B-CRACK-Q3_K_M.gguf 4.31 GB ee41d5e7 download
Q2_K 1 file 3.56 GB
Ornith-1.5-9B-CRACK-Q2_K.gguf 3.56 GB 981a5ec7 download
F16 1 file 879 MB
mmproj-Ornith-1.5-9B-f16.gguf 879 MB 90f5510a download
Auxiliary files 4 files 24.3 KB
dealign_mascot.png 10.9 KB da3bf39a download
dealign_logo.png 7.48 KB a5b3546b download
README.md 3.92 KB 1e973020 download
.gitattributes 1.94 KB 76a1959a download

README current version from Hugging Face


license: mit
library_name: gguf
pipeline_tag: image-text-to-text
base_model: ornith-ai/Ornith-1.5-9B
base_model_relation: quantized
tags:

  • gguf
  • llama.cpp
  • ornith
  • qwen3.5
  • abliterated
  • uncensored
  • crack
  • reasoning
  • vision
  • vlm

Dealign.ai
Dealign.ai

Ornith-1.5-9B-CRACK-GGUF

CRACK-abliterated Ornith 1.5 9B — GGUF quants for llama.cpp. Four quantizations
(Q8_0 / Q6_K / Q4_K_M / Q2_K) in one repository. Refusal behavior removed while preserving
the model's knowledge, reasoning ("thinking"), and full Vision-Language capability.

Ornith 1.5 is a hybrid GatedDeltaNet (SSM) + attention architecture; CRACK uses
architecture-aware weight surgery targeting the attention pathways, so knowledge and
coherence are retained (MMLU within ±3% of base at every quant).

Research artifact with reduced safety guardrails. Use responsibly and lawfully.

Quantizations

File Size Notes
Ornith-1.5-9B-CRACK-Q8_0.gguf 8.9 GB near-lossless reference
Ornith-1.5-9B-CRACK-Q6_K.gguf 7.4 GB near-lossless
Ornith-1.5-9B-CRACK-Q5_K_M.gguf 6.5 GB high quality
Ornith-1.5-9B-CRACK-Q4_K_M.gguf 5.6 GB balanced (recommended)
Ornith-1.5-9B-CRACK-Q3_K_M.gguf 4.6 GB small
Ornith-1.5-9B-CRACK-Q2_K.gguf 3.6 GB smallest

Pick one text file plus the vision projector mmproj-Ornith-1.5-9B-f16.gguf for image
input. Each quant is independently tuned (its own surgery strength) and verified — there
is no single strength shared across quants. Sub-8-bit quants use an AWQ (activation-aware)
pass plus an importance matrix for maximum quality.

Benchmarks

Evaluated through llama.cpp. MMLU is logit-mode accuracy (base vs. CRACK at the same
quant — isolates knowledge retention from quantization). HarmBench is coherence-gated
attack-success-rate over the 240 standard/contextual harm behaviors (copyright behaviors
excluded from the safety gate).

Quant MMLU (base) MMLU (CRACK) ΔMMLU HarmBench harm-ASR
Q8_0 78.1% 77.5% -0.53 pp 99.6%
Q6_K 76.5% 76.5% +0.00 pp 99.6%
Q5_K_M 76.5% 76.5% +0.00 pp 99.2%
Q4_K_M 78.3% 76.5% -1.76 pp 99.6%
Q3_K_M 73.3% 74.4% +1.06 pp 99.2%
Q2_K 50.5% 50.5% +0.00 pp 99.2%

MMLU is retained within ±3 pp of base at every quant. (Q2_K's absolute MMLU is lower because
2-bit quantization alone costs ~27 pp on a 9B — the surgery adds no further loss.)

HarmBench harm-ASR by topic (CRACK)

Topic harm-ASR
chemical / biological 100.0%
cybercrime / intrusion 100.0%
harassment / bullying 100.0%
harmful 100.0%
illegal 100.0%
misinformation / disinformation 98.1%

Usage (llama.cpp)

llama-cli -m Ornith-1.5-9B-CRACK-Q4_K_M.gguf -cnv --jinja \
  --temp 1.0 --top-p 0.95 --top-k 20
# or serve:
llama-server -m Ornith-1.5-9B-CRACK-Q4_K_M.gguf --jinja \
  --temp 1.0 --top-p 0.95 --top-k 20 -c 8192

Recommended sampling: temperature=1.0, top_p=0.95, top_k=20.

Reasoning

Ornith 1.5 emits a <think> reasoning trace and it is ON by default. To disable it, pass
{"chat_template_kwargs": {"enable_thinking": false}} to the chat endpoint. Works out of the
box in LM Studio.

Vision (image + text)

This is a multimodal model. Download a text quant and mmproj-Ornith-1.5-9B-f16.gguf:

llama-mtmd-cli -m Ornith-1.5-9B-CRACK-Q4_K_M.gguf \
  --mmproj mmproj-Ornith-1.5-9B-f16.gguf --jinja \
  --image photo.jpg -p "Describe this image."
# or serve with vision:
llama-server -m Ornith-1.5-9B-CRACK-Q4_K_M.gguf \
  --mmproj mmproj-Ornith-1.5-9B-f16.gguf --jinja -c 8192

The same mmproj works with all four text quants.

License

MIT (inherited from the upstream Ornith 1.5 base model).

Contact

[email protected]

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-24Duplicate from dealignai/Ornith-1.5-9B-UNCENSORED-GGUF02928a23.9 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration