← back to catalog · registered 2026-08-22 13:56

KYUNGYONG/aya-expanse-32b-abliterated-Q4-mlx

KYUNGYONG 32B second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/KYUNGYONG%2Faya-expanse-32b-abliterated-Q4-mlx"
Response includes
  • classification m1
  • files 11
  • hub_downloads_all_time 478
  • author_summary 13 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
478
22 last 30d - cooling
Likes
0
Model age
19mo ago
created 2025-03-05
Downloads over time
Now486→from13↑3,638%
017835653313 on Mar 5, 2025486 on Oct 11486 on Oct 10Mar '25Jun '25Sep '25Dec '25MarJunSep
Mar 5, 2025 → Oct 11 · 123 snapshots · spans 585 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

Languages
en fr de es it pt ja ko zh ar el fa pl id cs he hi nl ro ru tr uk vi
Tags
transformers safetensors cohere text-generation abliterated uncensored mlx mlx-my-repo conversational en fr de

Related

Total size
16.9 GB
Files
11
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2025-03-05 08:54

Files by quantization

Auxiliary files 11 files 16.9 GB
model-00003-of-00004.safetensors 4.96 GB b2c9d4bf download
model-00001-of-00004.safetensors 4.95 GB b1e64e97 download
model-00002-of-00004.safetensors 4.93 GB 93c1221a download
model-00004-of-00004.safetensors 2.08 GB 022c95ed download
tokenizer.json 19.2 MB 345ccf04 download
model.safetensors.index.json 73.6 KB 90df315f download
tokenizer_config.json 8.47 KB 024a6113 download
README.md 1.67 KB c8a88e32 download
.gitattributes 1.53 KB 52373fe2 download
config.json 842 B 47f75ba3 download
special_tokens_map.json 439 B cf33f2e5 download

README current version from Hugging Face


inference: false
library_name: transformers
language:

  • en
  • fr
  • de
  • es
  • it
  • pt
  • ja
  • ko
  • zh
  • ar
  • el
  • fa
  • pl
  • id
  • cs
  • he
  • hi
  • nl
  • ro
  • ru
  • tr
  • uk
  • vi
    license: cc-by-nc-4.0
    extra_gated_prompt: By submitting this form, you agree to the License Agreement and
    acknowledge that the information you provide will be collected, used, and shared
    in accordance with Cohere’s Privacy Policy. You’ll
    receive email updates about C4AI and Cohere research, events, products and services.
    You can unsubscribe at any time.
    extra_gated_fields:
    Name: text
    Affiliation: text
    Country: country
    I agree to use this model for non-commercial use ONLY: checkbox
    tags:
  • abliterated
  • uncensored
  • mlx
  • mlx-my-repo
    base_model: huihui-ai/aya-expanse-32b-abliterated

KYUNGYONG/aya-expanse-32b-abliterated-Q4-mlx

The Model KYUNGYONG/aya-expanse-32b-abliterated-Q4-mlx was converted to MLX format from huihui-ai/aya-expanse-32b-abliterated using mlx-lm version 0.21.5.

Use with mlx

pip install mlx-lm
from mlx_lm import load, generate

model, tokenizer = load("KYUNGYONG/aya-expanse-32b-abliterated-Q4-mlx")

prompt="hello"

if hasattr(tokenizer, "apply_chat_template") and tokenizer.chat_template is not None:
    messages = [{"role": "user", "content": prompt}]
    prompt = tokenizer.apply_chat_template(
        messages, tokenize=False, add_generation_prompt=True
    )

response = generate(model, tokenizer, prompt=prompt, verbose=True)

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2025-03-05Upload README.md with huggingface_hub34558e71.7 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration