← back to catalog · registered 2026-08-22 13:56

chimingw/qwen3.5-4b-uncensored-hauhaucs-aggressive-q6-k-llamafile

chimingw Qwen 4B GGUF multimodal second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/chimingw%2Fqwen3.5-4b-uncensored-hauhaucs-aggressive-q6-k-llamafile"
Response includes
  • classification m-uncensored
  • files 4
  • hub_downloads_all_time 48
  • author_summary 9 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
48
24 last 30d - active
Likes
0
Model age
8w ago
created 2026-08-15
Downloads over time
Now58→from14↑314%
1229466214 on Aug 1958 on Oct 1158 on Oct 8AugSepOct
Aug 19 → Oct 11 · 48 snapshots · spans 53 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en zh multilingual
Tags
llamafile gguf qwen3.5 multimodal vision uncensored quantized q6_k apple-silicon metal image-text-to-text en

Related

Total size
0 B
Files
4
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-15 03:49

Files by quantization

Auxiliary files 4 files 4.18 GB
Qwen3.5-4B-Uncensored-HauhauCS-Aggressive-Q6_K.llamafile 4.18 GB 54738b79 download
LICENSE 11.1 KB d6456956 download
README.md 5.01 KB b31e83a9 download
.gitattributes 1.57 KB c8db5c8c download

README current version from Hugging Face


license: apache-2.0
base_model: HauhauCS/Qwen3.5-4B-Uncensored-HauhauCS-Aggressive
base_model_relation: quantized
pipeline_tag: image-text-to-text
inference: false
language:

  • en
  • zh
  • multilingual
    tags:
  • llamafile
  • gguf
  • qwen3.5
  • multimodal
  • vision
  • uncensored
  • quantized
  • q6_k
  • apple-silicon
  • metal

Qwen3.5-4B Uncensored HauhauCS Aggressive Q6_K — llamafile

An unofficial, reproducible llamafile package of the Q6_K GGUF from HauhauCS/Qwen3.5-4B-Uncensored-HauhauCS-Aggressive.

HauHauCS claims:

0/465 refusals. Fully uncensored with zero capability loss.

No changes to datasets or capabilities. Fully functional, 100% of what the original authors intended — just without the refusals.

These are meant to be the best lossless uncensored models out there.

The single executable contains:

  • Qwen3.5-4B-Uncensored-HauhauCS-Aggressive-Q6_K.gguf;
  • mmproj-Qwen3.5-4B-Uncensored-HauhauCS-Aggressive-BF16.gguf for image input;
  • llamafile 0.10.5; and
  • terminal chat, browser chat, and a localhost API server.

The model and projector were embedded without retraining, merging, or re-quantization.

Intended use

This package is for people who want one downloadable executable containing a capable, uncensored, local Qwen3.5 4B model, its vision projector, a terminal chat interface, a browser chat interface, and a localhost API server. It is convenient for offline or self-contained macOS/Linux use where installing a separate inference runtime is undesirable.

Because everything is in one file, you can keep it on a USB drive and bring a self-contained local AI with you. That means you can carry a fully capable, uncensored AI on one drive—one that, according to HauHauCS's claim above, does not refuse prompts or instructions. The destination computer still needs a supported 64-bit platform and enough RAM, and the drive or filesystem must permit executable files.

Included artifact

File Size SHA-256
Qwen3.5-4B-Uncensored-HauhauCS-Aggressive-Q6_K.llamafile 4,490,544,382 bytes (4.49 GB / 4.18 GiB) 54738b795f32688a3951c5220357924f23b71899a39f227fe50c73eaea0d90a2

Embedded components

Component Size SHA-256
Qwen3.5-4B-Uncensored-HauhauCS-Aggressive-Q6_K.gguf 3,464,055,136 bytes ba93c21300854075ab42655bc30dca82c7c6c958f511d1ec9ea2b3e750b4b75f
mmproj-Qwen3.5-4B-Uncensored-HauhauCS-Aggressive-BF16.gguf 675,568,768 bytes a1e32e86ea99aa7a56f3dcfe7e63c1d0be9439d31fd07087099f15bc0fda0f22
llamafile-0.10.5 runtime 350,768,862 bytes 417bcc3348cd5162c2751812fc0ea2f6e79e89e7be6f17e8401ed95de2ed4246

Run

On macOS or Linux:

chmod +x Qwen3.5-4B-Uncensored-HauhauCS-Aggressive-Q6_K.llamafile
./Qwen3.5-4B-Uncensored-HauhauCS-Aggressive-Q6_K.llamafile

The terminal chat starts directly. While it is running, open the browser chat at:

http://127.0.0.1:8080/

The OpenAI-compatible API is available under:

http://127.0.0.1:8080/v1

The server is bound to localhost by default. Do not expose it publicly without authentication, TLS, and appropriate network controls.

Packaged defaults

  • context: 8,192 tokens;
  • sampler: temperature 0.6, top-p 0.95, top-k 20, min-p 0.0;
  • Jinja chat templates enabled;
  • multimodal image input enabled through the embedded BF16 projector.

To override the context size, pass --ctx-size when launching; for example:

./Qwen3.5-4B-Uncensored-HauhauCS-Aggressive-Q6_K.llamafile --ctx-size 32768

Larger contexts increase memory use. The upstream model advertises a much larger native context, but practical limits depend on the runtime, KV-cache settings, and available memory.

Provenance

Notes

  • This is an unofficial repackaging, not a new model release.
  • The embedded projector makes the package multimodal; image handling still depends on client/UI support in the embedded runtime.
  • The upstream repository describes the model as aggressively uncensored. This package does not independently reproduce or validate its refusal-rate claims.
  • The packaged 8,192-token default is conservative. Override it with --ctx-size N, such as --ctx-size 32768, if your memory budget permits.
  • This full-runtime artifact is larger than 4 GiB, so it is intended primarily for macOS and Linux rather than direct execution on Windows.
  • Hugging Face hosted inference is disabled because this repository distributes a standalone executable rather than a standard hosted-inference checkpoint.

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-15Fix HauhauCS claim quotation formatting08207255 KB
    Loading...
  2. 2026-08-15Add model README0879d1d5 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration