← back to catalog · registered 2026-08-22 13:56

Wondernutts/Ornith-1.0-35B-AEON-Ultimate-Uncensored-BF16-int4-ov

Wondernutts 35B MoE second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Wondernutts%2FOrnith-1.0-35B-AEON-Ultimate-Uncensored-BF16-int4-ov"
Response includes
  • classification m-uncensored
  • files 22
  • hub_downloads_all_time 139
  • author_summary 7 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
139
28 last 30d - stable
Likes
1
Model age
3mo ago
created 2026-07-04
Downloads over time
Now146→from0↑0%
0541071610 on Jul 1146 on Oct 11146 on Oct 9JulAugSepOct
Jul 1 → Oct 11 · 54 snapshots · spans 102 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en
Tags
openvino qwen3_5_moe int4 intel-arc qwen3_5 mixture-of-experts roleplay uncensored en base_model:AEON-7/Ornith-1.0-35B-AEON-Ultimate-Uncensored-BF16 base_model:finetune:AEON-7/Ornith-1.0-35B-AEON-Ultimate-Uncensored-BF16 license:apache-2.0

Related

Total size
18.3 GB
Files
22
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-07-08 10:06

Files by quantization

Auxiliary files 22 files 18.3 GB
openvino_language_model.bin 17.4 GB 126b02c8 download
openvino_text_embeddings_model.bin 485 MB 19a52947 download
openvino_vision_embeddings_merger_model.bin 423 MB a995b0d1 download
openvino_tokenizer.bin 9.18 MB b829798d download
openvino_detokenizer.bin 3.65 MB 48fc2c6a download
openvino_vision_embeddings_pos_model.bin 2.54 MB 261db08f download
openvino_vision_embeddings_model.bin 1.69 MB 22d5ef78 download
tokenizer.json 19.1 MB 6f32ce20 download
openvino_language_model.xml 7.51 MB c0ae136e download
openvino_vision_embeddings_merger_model.xml 1.26 MB e2229a9c download
openvino_tokenizer.xml 31.6 KB 993197ed download
openvino_detokenizer.xml 15.0 KB 8119bd04 download
openvino_vision_embeddings_model.xml 8.44 KB 9316ec9d download
chat_template.jinja 7.36 KB b07660cc download
openvino_text_embeddings_model.xml 5.76 KB 7ab30257 download
openvino_vision_embeddings_pos_model.xml 5.74 KB 1ee2390f download
README.md 3.67 KB e010cbbf download
config.json 3.11 KB d100b986 download
.gitattributes 1.53 KB 52373fe2 download
tokenizer_config.json 1.24 KB 9de16b5b download
openvino_config.json 1.18 KB 31f73387 download
generation_config.json 213 B a0d4001b download

README current version from Hugging Face


license: apache-2.0
base_model: AEON-7/Ornith-1.0-35B-AEON-Ultimate-Uncensored-BF16
tags:

  • openvino
  • int4
  • intel-arc
  • qwen3_5
  • mixture-of-experts
  • roleplay
  • uncensored
    language:
  • en

Ornith-1.0-35B (Qwen3.6-35B-A3B), OpenVINO INT4. 106 tok/s on ONE Intel Arc B70.

This is AEON-7's Ornith "Ultimate Uncensored" tune of Qwen3.6-35B-A3B (256-expert MoE, 3B
active parameters), converted to OpenVINO INT4 with Intel's own recipe. Sibling of my
tvall43 heretic conversion
of the same base model: same speed class, different tuning flavor. Runs on the stock 2026.2
runtime, no patches, coherent through 32K context.

One thing made this conversion special. Ornith was saved by a bleeding-edge transformers
(5.13-dev) that stores each of the 10,240 experts as separate tensors (30,720 tensor files
worth of them), while every current export toolchain expects the fused layout and silently
RANDOMIZES the experts if it does not find it. This conversion required re-fusing all 30,720
per-expert tensors back into the 80 fused ones, with the layout and packing order verified
against the toolchain's own model skeleton before writing. If you are trying to convert any
model saved in the new per-expert format, the fusion script and the full story are in the
toolkit repo. But, you should know
I did all that work to find out it was pretty tarded for roleplay, still good for workflow,
but not the best at campfire stories.

Measured performance (single Arc Pro B70, OpenVINO 2026.2)

Metric Value
Decode, short context ~106 tok/s
Decode at 6K context ~77 tok/s
Prefill pp512 ~1,100 tok/s; ~3,900 at 6K; 16K in ~7 s; 32K in ~22 s (curve measured on the tvall43 twin, same architecture)
Model load ~16 s
Weights ~19 GB (needs a 24 GB+ card; B60/B70 class)

Long-context capability, verified by needle retrieval

A password fact planted early in the prompt, retrieved at the end. Pass means the exact password.

Context Result
8K PASS (2.5 s total)
16K PASS (7.7 s)
24K PASS (19.0 s)
32K PASS (41.7 s, cold cache)

No rope patch, no precision workarounds, no special properties for single-stream use. The
architecture is rated to 262K positions; 32K is as far as my test box's host RAM lets me verify.

How to run

Identical to the tvall43 sibling:
standard Qwen <|im_start|> format, pre-closed <think>\n\n</think> block for fast no-think
replies (remove it and give at least 1024 tokens to enable reasoning), and
{"DYNAMIC_QUANTIZATION_GROUP_SIZE": 0} if you use continuous batching. Full runnable snippet
on the sibling card; just point it at this folder.

pip install openvino-genai==2026.2.0 huggingface_hub
huggingface-cli download Wondernutts/Ornith-1.0-35B-AEON-Ultimate-Uncensored-BF16-int4-ov --local-dir ./ornith-35b-ov

Provenance

Qwen/Qwen3.6-35B-A3B, Ornith-1.0 "AEON Ultimate Uncensored" tune by
AEON-7,
per-expert-to-fused weight restoration and OpenVINO INT4 conversion (this repo).

Intended use and content notice

Uncensored general model, built and tested for roleplay and creative writing on local Intel
hardware. The tune removes refusal behavior and outputs are unfiltered; you are responsible for
lawful and appropriate use. Licensed under
Apache 2.0, same as the upstream Qwen3.6 release.

README history 4 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-07-08Update README.mdea7e0a03.7 KB
    Loading...
  2. 2026-07-06Upload README.md with huggingface_hub8cf758f3.5 KB
    Loading...
  3. 2026-07-06Upload README.md with huggingface_huba2352de3.5 KB
    Loading...
  4. 2026-07-04Upload README.md with huggingface_hub9209ae43.4 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration