← back to catalog · registered 2026-08-27 02:02

ArtomYuan/Ornith-1.5-35B-A3B-abliterated-ROCmFPX

ArtomYuan 35B GGUF MoE multimodal second-order 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/ArtomYuan%2FOrnith-1.5-35B-A3B-abliterated-ROCmFPX"
Response includes
  • classification m8
  • files 5
  • hub_downloads_all_time 1,232
  • author_summary 4 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
1K
Likes
1
Model age
6w ago
created 2026-08-26
Downloads over time
Now4.7K→from128↑3,576%
01.7K3.5K5.2K128 on Aug 264.7K on Oct 11AugSepOct
Aug 26 → Oct 11 · 47 snapshots · spans 46 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
mit
Tags
gguf qwen35moe moe abliterated uncensored rocmfp4 rocmfpx llama.cpp-rocm vision multimodal base_model:huihui-ai/Huihui-Ornith-1.5-35B-A3B-abliterated base_model:quantized:huihui-ai/Huihui-Ornith-1.5-35B-A3B-abliterated

Related

Total size
17.7 GB
Files
5
Quantizations
2
Registered
2026-08-27 02:02
Last updated on HF
2026-10-08 02:23

Files by quantization

BF16 1 file 861 MB
mmproj-Ornith-1.5-35B-BF16.gguf 861 MB d9ce3102 download
Auxiliary files 4 files 17.7 GB
Ornith-1.5-35B-A3B-abliterated-Q4_ROCmFPX_FAST.gguf 17.7 GB 9b81dff7 download
README.md 5.23 KB 01121e43 download
.gitattributes 1.82 KB 173a92df download
sha256.txt 426 B 449ab552 download

README current version from Hugging Face


license: mit
base_model: huihui-ai/Huihui-Ornith-1.5-35B-A3B-abliterated
tags:

  • qwen35moe
  • abliterated
  • uncensored
  • rocmfp4
  • halofpx
  • gguf
  • vision
    library_name: gguf

Ornith-1.5-35B-A3B-abliterated-Q4_ROCmFPX_FAST

ROCmFPX quantized version of huihui-ai/Huihui-Ornith-1.5-35B-A3B-abliterated (locally requantized).

  • Base model: Ornith-1.5-35B-A3B (Claude-style 35B MoE, 3B active) abliterated (refusal behavior removed, uncensored)
  • Quantization: Q8_0 → ROCmFPX (Q4_0_ROCMFP4_FAST, 4.25 bpw), --allow-requantize
  • Vision: ✅ bundled mmproj (clip + qwen3vl_merger)

⚠ Important: Format Notice

ROCmFPX is a proprietary quantization format of the halofpx (ROCmFPX) engine — upstream llama.cpp CANNOT load this file!

Load it with halofpx (register this GGUF in the halofpx registry, then POST /api/v1/load).

Files

File Size sha256
Ornith-1.5-35B-A3B-abliterated-Q4_ROCmFPX_FAST.gguf ~17.6 GiB see sha256.txt
mmproj-Ornith-1.5-35B-BF16.gguf ~861 MiB see sha256.txt

Quantization Benchmarks

Measured on AMD Strix Halo (gfx1151), llama-bench, -p 256 -n 256 -t 16 -fa on -ngl 99, same prompt per row.

Variant tg256 (t/s) pp256 (t/s) Size bpw
Ornith-1.5-35B-A3B-abliterated-Q4_ROCMFP4_FAST (this repo) 73.0 829.7 17.6 GiB 4.25
Ornith-1.5-35B-A3B-abliterated-Q4_ROCMFP4 (base) 37.8 661.7 21.7 GiB 4.50
Ornith-1.5-35B-A3B-abliterated-Q8_ROCMFPX 47.6 813.5 34.2 GiB 8.0
Ornith-1.5-35B-A3B-abliterated-Q4_ROCMFP4_STRIX_LEAN 42.3 719.2 17.7 GiB 4.27

Why only Q4_ROCmFPX_FAST is published (other variants not uploaded):

  • Ornith-1.5-35B-A3B-abliterated-Q4_ROCMFP4 (base, dual-scale): 48% slower than FAST with only marginal precision gain — rejected
  • Ornith-1.5-35B-A3B-abliterated-Q4_ROCMFP4_STRIX_LEAN: 72% slower than FAST — rejected
  • Ornith-1.5-35B-A3B-abliterated-Q8_ROCMFPX: near-lossless but 2× size and 35% slower — to be published later as the quality tier

Usage

# After registering in halofpx registry (example):
curl -X POST http://127.0.0.1:8010/api/v1/load \
  -H "Authorization: Bearer ***" \
  -d '{"model_id":"ornith-1.5-35b-abliterated","reasoning_mode":"off"}'
  • Reasoning model: reasoning_mode: off outputs directly; MTP is a net loss for this model, keep it off
  • Context: up to 256K supported (measured same speed as 131K)

Disclaimer

This is an abliterated (uncensored) version and may produce outputs that do not conform to safety policies. Use at your own risk. No warranty is provided.


Ornith-1.5-35B-A3B-abliterated-Q4_ROCmFPX_FAST(ROCmFPX)

基于 huihui-ai/Huihui-Ornith-1.5-35B-A3B-abliterated 的 ROCmFPX 量化版(本地 requantize 产物)。

  • 原模型:Ornith-1.5-35B-A3B(Claude 风格 35B MoE,3B 激活)的 abliterated 消融版(去拒绝行为,无审查)
  • 量化:Q8_0 → ROCmFPX(Q4_0_ROCMFP4_FAST,4.25 bpw),--allow-requantize
  • 视觉:✅ 附带 mmproj(clip + qwen3vl_merger)

⚠ 重要:格式说明

ROCmFPX 是 halofpx(ROCmFPX)引擎专属量化格式——上游 llama.cpp 无法加载此文件!

请使用 halofpx 加载(halofpx registry 添加此 GGUF 后 POST /api/v1/load)。

文件

文件 大小 sha256
Ornith-1.5-35B-A3B-abliterated-Q4_ROCmFPX_FAST.gguf ~17.6 GiB 见仓库 sha256.txt
mmproj-Ornith-1.5-35B-BF16.gguf ~861 MiB 见仓库 sha256.txt

量化基准测试

测试环境:AMD Strix Halo(gfx1151),llama-bench,-p 256 -n 256 -t 16 -fa on -ngl 99,每行同一 prompt。

变体 tg256 (t/s) pp256 (t/s) 大小 bpw
Ornith-1.5-35B-A3B-abliterated-Q4_ROCMFP4_FAST(本仓库) 73.0 829.7 17.6 GiB 4.25
Ornith-1.5-35B-A3B-abliterated-Q4_ROCMFP4(基础版) 37.8 661.7 21.7 GiB 4.50
Ornith-1.5-35B-A3B-abliterated-Q8_ROCMFPX 47.6 813.5 34.2 GiB 8.0
Ornith-1.5-35B-A3B-abliterated-Q4_ROCMFP4_STRIX_LEAN 42.3 719.2 17.7 GiB 4.27

仅发布 Q4_ROCmFPX_FAST 的原因(其他量化不上传):

  • Ornith-1.5-35B-A3B-abliterated-Q4_ROCMFP4(基础版,双 scale):比 FAST 慢 48%,精度提升有限——弃用
  • Ornith-1.5-35B-A3B-abliterated-Q4_ROCMFP4_STRIX_LEAN:比 FAST 慢 72%——弃用
  • Ornith-1.5-35B-A3B-abliterated-Q8_ROCMFPX:近无损但体积 2 倍、慢 35%——待后续上传(高质量档)

使用

# halofpx registry 注册后(例):
curl -X POST http://127.0.0.1:8010/api/v1/load \
  -H "Authorization: Bearer ***" \
  -d '{"model_id":"ornith-1.5-35b-abliterated","reasoning_mode":"off"}'
  • 思考模型:reasoning_mode: off 直接输出;MTP 对该模型为净亏损,勿开
  • 上下文:支持 256K(实测与 131K 速度持平)

免责声明

本模型为 abliterated(消融去审查)版本,可能生成不符合安全政策的输出。使用者自行承担全部责任。本仓库不提供任何保证。

README history 19 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-10-07upload README.mdebb960b5.9 KB
    Loading...
  2. 2026-10-05docs: fix MTP enable flagf2074cb4 KB
    Loading...
  3. 2026-10-05EN-only README (unified template)2a18c373.9 KB
    Loading...
  4. 2026-09-21rename gguf files to model_quant namingd105c5e6.9 KB
    Loading...
  5. 2026-09-06docs: fix MTP support (llama.cpp-rocm supports --mtp) and clarify full upstre...104b4ed6.9 KB
    Loading...
  6. 2026-09-06docs: refine usage (--mmproj for vision, drop --special; align context note)ead3e0d6.7 KB
    Loading...
  7. 2026-09-06docs: recommend llama.cpp-rocm engine instead of halofpxa3ce9fe6.7 KB
    Loading...
  8. 2026-08-27Upload README.md with huggingface_hub9d19da76.7 KB
    Loading...
  9. 2026-08-27Upload README.md with huggingface_hub0339b126.6 KB
    Loading...
  10. 2026-08-27Upload README.md with huggingface_hub3cf899d5.5 KB
    Loading...
  11. 2026-08-27Upload README.md with huggingface_hubeec2b815.2 KB
    Loading...
  12. 2026-08-27Upload README.md with huggingface_hub48f1d7d5.2 KB
    Loading...
  13. 2026-08-27Upload README.md with huggingface_huba8864045.2 KB
    Loading...
  14. 2026-08-27Upload folder using huggingface_hub1050bb15.2 KB
    Loading...
  15. 2026-08-27Upload README.md with huggingface_hub90d04d05.2 KB
    Loading...
  16. 2026-08-27Upload folder using huggingface_hub5d964145.2 KB
    Loading...
  17. 2026-08-27Upload README.md with huggingface_hub80d19753.2 KB
    Loading...
  18. 2026-08-26Upload README.md with huggingface_hubfb2592c3.3 KB
    Loading...
  19. 2026-08-26Upload folder using huggingface_hub41fce842 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration