← back to catalog · registered 2026-08-22 13:56

mlx-community/Josiefied-Qwen2.5-Coder-7B-Instruct-abliterated-v1-4bit

mlx-community Qwen 7.6B GGUF second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/mlx-community%2FJosiefied-Qwen2.5-Coder-7B-Instruct-abliterated-v1-4bit"
Response includes
  • classification m8
  • files 11
  • hub_downloads_all_time 1,622
  • author_summary 207 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 3 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • author=mlx-community (M8 quantization producer)
  • is_gguf=1
  • base_model='Goekdeniz-Guelmez/Josiefied-Qwen2.5-Coder-7B-Instruct-abliterated-v1' looks abliterated -> assume M1
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
2K
240 last 30d - stable
Likes
1
Model age
20mo ago
created 2025-02-08
Downloads over time
Now1.7K→from9↑19,056%
06321.3K1.9K9 on Feb 5, 20251.7K on Oct 11Feb '25May '25Aug '25Nov '25FebMayAug
Feb 5, 2025 → Oct 11 · 127 snapshots · spans 613 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en de
Tags
mlx safetensors qwen2 chat GGUF text-generation conversational en de base_model:Goekdeniz-Guelmez/Josiefied-Qwen2.5-Coder-7B-Instruct-abliterated-v1 base_model:quantized:Goekdeniz-Guelmez/Josiefied-Qwen2.5-Coder-7B-Instruct-abliterated-v1 license:apache-2.0

Related

Total size
3.99 GB
Files
11
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2025-02-16 14:06

Files by quantization

Auxiliary files 11 files 4.00 GB
model.safetensors 3.99 GB 99e0d0b1 download
tokenizer.json 10.9 MB 63a2951d download
vocab.json 2.65 MB 4783fe10 download
merges.txt 1.59 MB 31349551 download
model.safetensors.index.json 50.5 KB 669ba50d download
tokenizer_config.json 10.6 KB 97cac88e download
.gitattributes 1.53 KB 52373fe2 download
README.md 1.11 KB 05c9c602 download
config.json 891 B 80885d36 download
special_tokens_map.json 610 B afc6dc69 download
added_tokens.json 605 B 482ced46 download

README current version from Hugging Face


language:

  • en
  • de
    license: apache-2.0
    tags:
  • chat
  • GGUF
  • mlx
    base_model: Goekdeniz-Guelmez/Josiefied-Qwen2.5-Coder-7B-Instruct-abliterated-v1
    pipeline_tag: text-generation

mlx-community/Josiefied-Qwen2.5-Coder-7B-Instruct-abliterated-v1-4bit

The Model mlx-community/Josiefied-Qwen2.5-Coder-7B-Instruct-abliterated-v1-4bit was
converted to MLX format from Goekdeniz-Guelmez/Josiefied-Qwen2.5-Coder-7B-Instruct-abliterated-v1
using mlx-lm version 0.21.3.

Use with mlx

pip install mlx-lm
from mlx_lm import load, generate

model, tokenizer = load("mlx-community/Josiefied-Qwen2.5-Coder-7B-Instruct-abliterated-v1-4bit")

prompt = "hello"

if tokenizer.chat_template is not None:
    messages = [{"role": "user", "content": prompt}]
    prompt = tokenizer.apply_chat_template(
        messages, add_generation_prompt=True
    )

response = generate(model, tokenizer, prompt=prompt, verbose=True)

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2025-02-08Add files using upload-large-folder tool6cb0e831.1 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration