← back to catalog · registered 2026-08-22 13:56

Shiftedx/ornith-1.0-35b-abliterated-mxfp4-vision-mtplx

Shiftedx 35B MoE multimodal
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Shiftedx%2Fornith-1.0-35b-abliterated-mxfp4-vision-mtplx"
Response includes
  • classification m1
  • files 20
  • hub_downloads_all_time 1,849
  • author_summary 24 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
2K
233 last 30d - stable
Likes
2
Model age
3mo ago
created 2026-07-05
Downloads over time
Now1.9K→from608↑220%
5411.1K1.6K2.1K608 on Jul 151.9K on Oct 11JulAugSepOct
Jul 15 → Oct 11 · 53 snapshots · spans 88 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Tags
mlx safetensors qwen3_5_moe mtplx mxfp4 apple-silicon local-inference privacy vision abliterated image-text-to-text conversational

Related

Total size
18.5 GB
Files
20
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-20 17:26

Files by quantization

Auxiliary files 20 files 18.6 GB
model-00003-of-00004.safetensors 5.00 GB 3989f453 download
model-00002-of-00004.safetensors 5.00 GB 1015e82b download
model-00001-of-00004.safetensors 4.98 GB 3dc6088c download
model-00004-of-00004.safetensors 2.18 GB 6531899b download
model-vision-00001-of-00001.safetensors 852 MB d4579339 download
mtp.safetensors 554 MB d9788ed6 download
tokenizer.json 19.1 MB 06b95093 download
vocab.json 6.41 MB 0aa0ce06 download
model.safetensors.index.json 161 KB f1b02b90 download
config.json 24.2 KB 4ccdd41c download
tokenizer_config.json 8.81 KB 1a5e6600 download
chat_template.jinja 7.36 KB 33dc1028 download
README.md 3.29 KB d8742c69 download
mtplx_runtime.json 2.51 KB cf8ef438 download
.gitattributes 1.53 KB 52373fe2 download
processor_config.json 991 B 8f29fe38 download
MTPLX_NOTES.md 435 B ea63281c download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
generation_config.json 213 B 3f7527c4 download

README current version from Hugging Face


license: other
base_model: deepreinforce-ai/Ornith-1.0-35B
library_name: mlx
pipeline_tag: image-text-to-text
tags:

  • mlx
  • mtplx
  • mxfp4
  • qwen3_5_moe
  • apple-silicon
  • local-inference
  • privacy
  • vision
  • abliterated

Ornith 1.0 35B Abliterated MXFP4 Vision MTPLX

Vision-enabled MXFP4 MLX build of deepreinforce-ai/Ornith-1.0-35B, packaged for MTPLX native-MTP inference on Apple Silicon.

This is intended for local, private inference. The package contains model files only: no hosted endpoint, telemetry, prompt logs, or external service requirement.

Notes

  • Includes the Ornith vision tower and processor files.
  • Optimized for MTPLX MTP serving, not LM Studio indexing.
  • Uses a compatible prequantized q5/g64 MTP sidecar; recommended draft depth is 2.
  • Chat template defaults to thinking off unless enable_thinking=true is passed explicitly.
  • Abliteration metadata is included for transparency; no source direction file is required for inference.

Local Validation

Hardware reference: Apple M4 Max Apple Silicon with 64 GB unified memory.

Check Result
API health Pass
Text JSON smoke Pass
Executable code smoke 3/3
Vision smoke Pass
Mean decode speed 158.0 tok/s
Accepted draft ratio 92.3%

The image smoke identified HUNTER and a living room scene. These are lightweight local checks, not public leaderboard scores.

Recommended MTPLX Settings

Use depth 2, profile sustained, tokenizer chat template, MTP enabled, and thinking disabled by default.

Shiftedx Bench post-publication qualification

This table was generated from the frozen lightweight quant gate after the model weights were published. Categories remain separate; the benchmark does not produce a composite intelligence score.

Lane Passed Accuracy Mean wall time Mean decode Peak active memory
Quality 5/10 50.0% 9.60 s 160.47 tok/s 35.28 GiB
Long context 1/15 6.7% 36.35 s 131.56 tok/s 37.96 GiB
Tool calling 5/6 83.3% 1.32 s 131.81 tok/s 36.45 GiB
Agentic 1/2 50.0% 3.30 s — tok/s —
Vision 0/4 0.0% 1.72 s 165.93 tok/s 35.15 GiB
  • Tested model revision: afed7f14a357c9db38bf67eb5b7c66b6152157f5
  • Benchmark: Shiftedx Bench v0.3.0
  • Context lengths represented: 4,096, 16,384, 65,536, 131,072 prompt tokens; effective tested context: not established at the benchmark's 90% threshold
  • Runtime contract: MTPLX 2.7.1 sustained; thinking on/medium; sampler temperature=0.2, top_p=0.95, top_k=20; Hermes hybrid routing with native explicit-parallel calls; KV cache off; MTP depth 2
  • Host: Apple M4 Max, 64 GiB unified memory
  • Total measured request wall time: 662.69 seconds
  • 260,096-token status: not run; it is outside the lightweight quant gate.

Scores are specific to the linked model revision, benchmark revision, runtime contract, and host. Changing weight precision, KV-cache precision, reasoning mode, template, or speculative depth creates a different benchmark candidate.

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-17docs: add standardized Shiftedx quant gateff4e4903.3 KB
    Loading...
  2. 2026-07-05Add validation hardware referenceafed7f11.5 KB
    Loading...
  3. 2026-07-05Upload abliterated Ornith MXFP4 MTPLX vision build351bd221.4 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration