← back to catalog · registered 2026-10-02 00:58

RaiseRuntimeError/Huihui-Qwen3.5-2B-abliterated-q4f16_1-MLC

RaiseRuntimeError 2B second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/RaiseRuntimeError%2FHuihui-Qwen3.5-2B-abliterated-q4f16_1-MLC"
Response includes
  • classification m-uncensored
  • files 40
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
0
Likes
0
Model age
today
created 2026-10-02

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Languages
en
Tags
mlc-llm web-llm webgpu qwen3.5 abliterated uncensored q4f16_1 text-generation conversational en base_model:huihui-ai/Huihui-Qwen3.5-2B-abliterated base_model:quantized:huihui-ai/Huihui-Qwen3.5-2B-abliterated

Related

Total size
1010 MB
Files
40
Quantizations
1
Registered
2026-10-02 00:58
Last updated on HF
2026-10-02 00:28

Files by quantization

Auxiliary files 40 files 1.01 GB
params_shard_0.bin 243 MB 30f8dbb0 download
params_shard_26.bin 31.5 MB 0541d711 download
params_shard_27.bin 31.5 MB ac12ed68 download
params_shard_28.bin 31.5 MB 5e07540d download
params_shard_29.bin 31.5 MB 32e9676e download
params_shard_1.bin 30.3 MB 5e81f1b7 download
params_shard_10.bin 27.0 MB 290b95b6 download
params_shard_12.bin 27.0 MB 4d6ba09e download
params_shard_14.bin 27.0 MB 2684b667 download
params_shard_16.bin 27.0 MB e929256c download
params_shard_18.bin 27.0 MB e3524499 download
params_shard_19.bin 27.0 MB ee4e77d9 download
params_shard_2.bin 27.0 MB b0dbb7a3 download
params_shard_20.bin 27.0 MB 49feae9b download
params_shard_21.bin 27.0 MB 0cd01377 download
params_shard_22.bin 27.0 MB bed9ca6b download
params_shard_23.bin 27.0 MB 5f29044e download
params_shard_25.bin 27.0 MB 09760cd3 download
params_shard_7.bin 27.0 MB 7f8110ad download
params_shard_8.bin 27.0 MB 225ea2f9 download
params_shard_9.bin 27.0 MB 13a7565c download
params_shard_30.bin 24.2 MB 8985f3e8 download
params_shard_11.bin 20.3 MB fcfd29ed download
params_shard_13.bin 20.3 MB 4543a367 download
params_shard_15.bin 20.3 MB 21de7864 download
params_shard_17.bin 20.3 MB e7d17a85 download
params_shard_24.bin 20.3 MB 81fff5d8 download
params_shard_3.bin 20.3 MB 078a08b5 download
params_shard_4.bin 20.3 MB 3e90369c download
params_shard_5.bin 20.3 MB cc36dc82 download
params_shard_6.bin 20.3 MB ce4ea315 download
tokenizer.json 12.2 MB 5f9e4d49 download
vocab.json 6.41 MB 0aa0ce06 download
merges.txt 3.20 MB a494e019 download
tensor-cache.json 168 KB e465cd35 download
tensor-cache-b16.json 168 KB 27ef54bd download
tokenizer_config.json 16.3 KB fae3ce99 download
mlc-chat-config.json 2.30 KB ed77c801 download
README.md 1.63 KB b31ee429 download
.gitattributes 1.53 KB 52373fe2 download

README current version from Hugging Face


language:

  • en
    library_name: mlc-llm
    base_model: huihui-ai/Huihui-Qwen3.5-2B-abliterated
    license: other
    tags:
  • mlc-llm
  • web-llm
  • webgpu
  • qwen3.5
  • abliterated
  • uncensored
  • q4f16_1
    pipeline_tag: text-generation

Huihui Qwen3.5-2B (abliterated) — MLC q4f16_1

MLC-converted (q4f16_1) build of the abliterated
huihui-ai/Huihui-Qwen3.5-2B-abliterated
model, for running fully client-side in the browser via
WebLLM (WebGPU).

  • Base model: huihui-ai/Huihui-Qwen3.5-2B-abliterated (all credit to Huihui for the abliteration)
  • Quantization: q4f16_1
  • Architecture: Qwen3.5-2B (model_type: qwen3_5, hidden 2048, 24 layers)
  • Converted with: mlc-llm convert_weight (source build)

Usage (WebLLM)

import * as webllm from "https://esm.run/@mlc-ai/[email protected]";

const appConfig = {
  model_list: [
    {
      model: "https://huggingface.co/RaiseRuntimeError/Huihui-Qwen3.5-2B-abliterated-q4f16_1-MLC",
      model_id: "Huihui-Qwen3.5-2B-abliterated-q4f16_1-MLC",
      model_lib:
        webllm.modelLibURLPrefix + webllm.modelVersion + "/Qwen3.5-2B-q4f16_1_cs1k-webgpu.wasm",
      vram_required_MB: 2245.44,
      overrides: { context_window_size: 4096, max_history_size: 1 },
    },
  ],
};

const engine = await webllm.CreateMLCEngine(
  "Huihui-Qwen3.5-2B-abliterated-q4f16_1-MLC",
  { appConfig },
);

Notes

This is a derivative of a third-party abliterated model, re-quantized to MLC
q4f16_1 for browser inference. The abliteration itself is the work of
huihui-ai.

Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.