← back to catalog · registered 2026-08-22 13:56

Atomic-Germ/Huihui-Qwythos-9B-Claude-Mythos-5-1M-abliterated-NPU2

Atomic-Germ Qwen 9B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Atomic-Germ%2FHuihui-Qwythos-9B-Claude-Mythos-5-1M-abliterated-NPU2"
Response includes
  • classification m1
  • files 9
  • hub_downloads_all_time 1,763
  • author_summary 3 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
2K
161 last 30d - cooling
Likes
0
Model age
2mo ago
created 2026-08-07
Downloads over time
Now1.8K→from370↑397%
2978601.4K2K370 on Aug 51.8K on Oct 11AugSepOct
Aug 5 → Oct 11 · 50 snapshots · spans 67 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en
Tags
qwen3_5_text qwen3.5 qwythos 1M-context long-context agentic uncensored abliterated fastflowlm q4nx npu text-generation

Related

Total size
0 B
Files
9
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-19 20:07

Files by quantization

Auxiliary files 9 files 8.20 GB
model.q4nx 7.23 GB bc4608cd download
vision_weight.q4nx 980 MB 07979ede download
tokenizer.json 12.2 MB 5f9e4d49 download
flm-add.py 23.8 KB 68d8d145 download
tokenizer_config.json 14.8 KB 03559b06 download
chat_template.jinja 7.57 KB a585dec8 download
config.json 3.19 KB d1889bf7 download
README.md 2.96 KB 710f1c87 download
.gitattributes 2.15 KB a40d672b download

README current version from Hugging Face


license: apache-2.0
language:

  • en
    pipeline_tag: text-generation
    tags:
  • qwen3.5
  • qwythos
  • 1M-context
  • long-context
  • agentic
  • uncensored
  • abliterated
  • fastflowlm
  • q4nx
  • npu
    base_model:
  • empero-ai/Qwythos-9B-Claude-Mythos-5-1M
    base_model_relation: quantized
    quantized_by: Atomic-Germ

IF YOU USE COMMUNITY QWEN MODELS DO NOT UPGRADE TO FLM v1.0.2+

Qwythos-9B-Claude-Mythos-5-1M (abliterated) - Q4NX for FastFlowLM (AMD Ryzen AI XDNA2)

Qwythos-9B-Claude-Mythos-5-1M (abliterated) is a full fine-tune of Qwen3.5-9B for long-context agentic use (1M context), with abliterated safety tuning. This repository contains the Q4NX conversion for FastFlowLM.

What is Q4NX?

Q4NX is FastFlowLM's native packed-quantization format - a rearranged Q4_1
layout tuned for the NPU matrix engine's tile sizes and memory access
patterns. It is not a GGUF file and it does not run on llama.cpp or
Ollama; it is meant exclusively for the FastFlowLM
engine on AMD Ryzen AI NPUs.

Requirements

  • FastFlowLM >= 0.9.45 (flm CLI)
  • AMD Ryzen AI processor with XDNA2 (NPU2) - Strix Point / Ryzen AI 300
    series or later
  • Linux with the XRT NPU stack installed
  • ~17 GB of unified system memory (Q4NX weights + activations + KV cache)

Files

File Purpose
model.q4nx Quantized Q4NX weights
config.json FastFlowLM model configuration
tokenizer.json Tokenizer
tokenizer_config.json Special tokens and chat template
chat_template.jinja Chat template (optional)
flm-add.py Installer script - registers this model with FastFlowLM

Install and run

This repository works with flm-add, a small installer that copies the model
into the FastFlowLM user directory and registers the tag. It never
modifies the system FastFlowLM install.

pip install flm-add or uv tool install flm-add

uv tool install flm-add
flm-add Atomic-Germ/Huihui-Qwythos-9B-Claude-Mythos-5-1M-abliterated-NPU2 --tag qwythos:9b --family qwen3.5
FLM_CONFIG_PATH="$HOME/.config/flm/model_list.json" FLM_XCLBIN_PATH="$HOME/.config/flm" flm run qwythos:9b

Kernels

FastFlowLM's NPU kernels (xclbins) are closed source and are not shipped in
this repository. flm-add.py links the kernels of the official qwen3.5:9b
model (Qwen3.5-9B-NPU2), because this model shares the same engine family
(qwen3.5) and architecture.

Model

  • Registry tag: qwythos:9b
  • Engine family: qwen3.5
  • Kernel source: Qwen3.5-9B-NPU2
  • Context length: 262,144 tokens (from config)
  • model.q4nx size: 7.76 GB
  • Base model: empero-ai/Qwythos-9B-Claude-Mythos-5-1M
  • License: apache-2.0

Original model card

See the upstream model card for training details, benchmarks, and upstream
usage. This repository only contains the Q4NX conversion for FastFlowLM.

README history 6 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-19Update README.mde043d473 KB
    Loading...
  2. 2026-08-16Update README.md5d455872.9 KB
    Loading...
  3. 2026-08-14flm-add available from pypi3cddfbd2.8 KB
    Loading...
  4. 2026-08-08Q4NX/FastFlowLM card: flm-add.py install + run/serve flow, kernels source, mo...903cd223.5 KB
    Loading...
  5. 2026-08-07Update README.md42f2a1e2.4 KB
    Loading...
  6. 2026-08-07Duplicate from huihui-ai/Huihui-Qwythos-9B-Claude-Mythos-5-1M-abliterated-GGUF6506e483.1 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration