← back to catalog · registered 2026-08-22 13:56

choppedgarlic/Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-4bit-MLX

choppedgarlic Qwen 27B second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/choppedgarlic%2FQwen3.8-27B-AEON-ULTIMATE-UNCENSORED-4bit-MLX"
Response includes
  • classification m-uncensored
  • files 12
  • hub_downloads_all_time 21,610
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
22K
3K last 30d - stable
Likes
11
Model age
8w ago
created 2026-08-15
Downloads over time
Now22.4K→from0↑0%
08.2K16.4K24.6K0 on Aug 1522.4K on Oct 11AugSepOct
Aug 15 → Oct 11 · 49 snapshots · spans 57 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en
Tags
mlx safetensors qwen3_5 apple-silicon quantized 4-bit qwen3.8 reasoning text-generation conversational en base_model:AEON-7/Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-BF16

Related

Total size
14.1 GB
Files
12
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-15 21:23

Files by quantization

Auxiliary files 12 files 14.1 GB
model-00002-of-00003.safetensors 4.99 GB 4cccbe3e download
model-00001-of-00003.safetensors 4.96 GB 227d9596 download
model-00003-of-00003.safetensors 4.14 GB d4ff7b5c download
tokenizer.json 19.1 MB 06b95093 download
model.safetensors.index.json 185 KB 8cabf784 download
LICENSE 11.1 KB d6456956 download
chat_template.jinja 8.74 KB c0c686f9 download
config.json 4.01 KB 7926803c download
README.md 3.11 KB a556a338 download
.gitattributes 1.53 KB 52373fe2 download
tokenizer_config.json 1.24 KB 0b282af8 download
generation_config.json 214 B 3f25ead4 download

README current version from Hugging Face


license: apache-2.0
language:

  • en
    pipeline_tag: text-generation
    library_name: mlx
    base_model: AEON-7/Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-BF16
    tags:
  • mlx
  • apple-silicon
  • quantized
  • 4-bit
  • qwen3.8
  • reasoning

Qwen3.8 27B AEON Ultimate Uncensored — 4-bit MLX

This is a 4-bit MLX quantization for Apple Silicon of
AEON-7/Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-BF16.

Original Uncensored model made by @SpaceTimeViking , complete credit to him and all his work
https://x.com/SpaceTimeViking

All model and fine-tuning credit belongs to the original author, AEON-7, and
the upstream Qwen team. This repository only provides an MLX conversion and
quantization for easier local use on Apple Silicon. No additional fine-tuning
or intentional behavioral changes were made.

Quantization

  • Format: MLX
  • Quantization: 4-bit affine
  • Group size: 64
  • Approximate download size: 14 GB
  • Intended platform: Apple Silicon macOS
  • Recommended unified memory: 32 GB or more

Install

pip install -U mlx-lm

Run an interactive chat

mlx_lm.chat \
  --model choppedgarlic/Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-4bit-MLX \
  --max-tokens 8192 \
  --temp 0.6 \
  --top-p 0.95

The first launch downloads the model from Hugging Face. Later launches use the
local Hugging Face cache.

Run an OpenAI-compatible local server

Thinking enabled:

mlx_lm.server \
  --model choppedgarlic/Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-4bit-MLX \
  --host 127.0.0.1 \
  --port 8081 \
  --max-tokens 8192 \
  --temp 0.6 \
  --top-p 0.95 \
  --chat-template-args '{"enable_thinking":true}'

Thinking disabled:

mlx_lm.server \
  --model choppedgarlic/Qwen3.8-27B-AEON-ULTIMATE-UNCENSORED-4bit-MLX \
  --host 127.0.0.1 \
  --port 8081 \
  --max-tokens 8192 \
  --temp 0.6 \
  --top-p 0.95 \
  --chat-template-args '{"enable_thinking":false}'

The API base URL is then http://127.0.0.1:8081/v1.

Verified locally

  • Text generation
  • Interactive mlx_lm.chat
  • Thinking and non-thinking modes
  • mlx_lm.server through its OpenAI-compatible API
  • Local use through the Pi coding agent

Limitations

  • This conversion is intended for Apple Silicon and requires MLX.
  • The current conversion contains no vision_config or vision preprocessor
    assets. Image input and image generation are not supported or claimed.
  • Quantization may reduce quality compared with the BF16 source model.
  • The model can produce inaccurate, unsafe, or objectionable output. Validate
    outputs before using them in production or high-stakes settings.
  • Please also read the original model card for its intended use, behavior, and
    limitations.

License and attribution

Released under the Apache License 2.0, matching the source repository's declared
license. See LICENSE.

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-15Update README.md875c8523.1 KB
    Loading...
  2. 2026-08-15Upload folder using huggingface_hubc1608c73 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration