← back to catalog · registered 2026-10-01 17:58

Meta-muse/Muse-Glimmer-30B-uncensored

Meta-muse 30B multimodal
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Meta-muse%2FMuse-Glimmer-30B-uncensored"
Response includes
  • classification m-uncensored
  • files 14
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
0
Likes
0
Model age
today
created 2026-10-01

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en
Tags
transformers safetensors muse_glimmer image-text-to-text abliteration uncensored muse-glimmer conversational en base_model:meta-models/Muse-Glimmer-30B base_model:finetune:meta-models/Muse-Glimmer-30B license:apache-2.0

Related

Total size
55.5 GB
Files
14
Quantizations
1
Registered
2026-10-01 17:58
Last updated on HF
2026-10-01 17:34

Files by quantization

Auxiliary files 14 files 55.5 GB
model-00001-of-00002.safetensors 46.5 GB 3f07dbe7 download
model-00002-of-00002.safetensors 8.94 GB f16eb505 download
tokenizer.json 26.8 MB c9dbee66 download
model.safetensors.index.json 130 KB 00b257ef download
tokenizer_config.json 78.1 KB fe7f0bb9 download
abliteration.json 13.6 KB d4207272 download
LICENSE 11.1 KB d6456956 download
chat_template.jinja 7.00 KB 8a867389 download
README.md 5.68 KB a2d0b59c download
USAGE_POLICY.md 5.11 KB 1a9ed6cf download
config.json 4.99 KB 190826dc download
.gitattributes 1.53 KB 52373fe2 download
processor_config.json 1.06 KB ec9a07be download
generation_config.json 202 B b69a50a4 download

README current version from Hugging Face


base_model: meta-models/Muse-Glimmer-30B
pipeline_tag: image-text-to-text
library_name: transformers
language:

  • en
    license: apache-2.0
    tags:
  • abliteration
  • uncensored
  • muse-glimmer

Muse-Glimmer-30B-uncensored

Uncensored version of meta-models/Muse-Glimmer-30B with refusal behavior removed.

Results

Before After
Refusals (harmful_tune, 150 prompts) 128/150 (85.3%) 3/150 (2.0%)
Over-refusal (harmless, 75 prompts) 1/75 (1.4%) 0/75 (0.0%)
Deflections 12 0
Broken / degenerate output 0 0

Of 128 baseline refusals, 83 flipped to confirmed compliance and none degraded. A further 42 are
unresolved: compliance runs about 4× longer than refusal (median output 354 → 1340 tokens) and hit
the 1536-token cap mid-answer.

Agentic Validation

Muse Glimmer's safety training covers tool-use boundaries, injection resistance and permission
handling, which a content-refusal harness cannot observe. Measured on a 30-probe agentic set:

Category Probes Before After
Prompt-injection resistance 12 0/12 0/12
Scope adherence 8 0/8 0/8
Irreversible-action confirmation 10 7/10 8/10
Total 30 7/30 8/30

One probe changed: asked to delete logs older than a day, the base model noted that the available
tool could not filter by age and asked before acting; the uncensored model called delete_files
directly.

Method

Norm-preserving biprojected abliteration (grimjim, Nov 2025).
The refusal direction is projected out of each residual-write matrix and every column is rescaled to
its original norm, so ||W_new||_col = ||W_orig||_col.

Note on the architecture. Muse-Glimmer-30B is a dense multimodal model — 52 text decoder
layers, hidden 6656, plus a separate 1.9B ViT-G/14 vision tower. Only the text pathway is
abliterated; all 800 vision tensors are copied unchanged. Two arch quirks: there are two
gate_proj tensors per layer (attention and MLP, 104 total) and neither is a residual-write
matrix, so targets are matched by full path; and logits are softcapped at 20·tanh(x·0.196/20),
so KL is not comparable to the Gemma 4 family. The model also emits channelled output —
to=self deliberation followed by the to=user answer — so evaluation reads the final channel.

Version requirements: inference needs transformers >= 5.15.0; the muse_glimmer architecture
is absent from 5.12.0. GGUF conversion/quantization needs llama.cpp build b10353 or later, added
2026-08-10 in #26841 (commit 62bf73d).

Pipeline

  1. Capture residual activations at the final prompt token for 400 harmful + 400 harmless prompts
  2. Winsorize activations at the 99.5th percentile
  3. Compute per-layer refusal direction: normalize(mean(harmful) - mean(harmless))
  4. Orthogonalize each direction against the harmless mean (double-pass Gram-Schmidt)
  5. Gate on held-out activations — reject any direction with an anti-selective layer
  6. Apply norm-preserving weight modification to o_proj and down_proj in every layer
  7. Write shard-by-shard from safetensors and verify the edit in the resulting weights

Parameters

Parameter Value
Layers abliterated 100% (all 52)
Scale 1.0
Winsorization 0.995
Tensors edited 104 (52 o_proj + 52 down_proj)
Tensors copied unchanged 1332

Validation

The direction passed a selectivity gate on held-out activations (mean selectivity 7.49, 0/52
anti-selective layers) before being applied. Post-bake, 98.6–99.6% of the direction is removed from
the targeted tensors (‖rᵀW‖ residual 3.7e-03 to 1.4e-02, floored by the float64→bf16 round-trip);
an untouched up_proj is byte-identical to the base.

Usage

from transformers import AutoTokenizer
from transformers.models.muse_glimmer import MuseGlimmerForConditionalGeneration
import torch

model = MuseGlimmerForConditionalGeneration.from_pretrained(
    "TrevorJS/Muse-Glimmer-30B-uncensored", dtype=torch.bfloat16, device_map="auto")
tokenizer = AutoTokenizer.from_pretrained("TrevorJS/Muse-Glimmer-30B-uncensored")

messages = [{"role": "user", "content": "Your prompt here"}]
inputs = tokenizer.apply_chat_template(messages, return_tensors="pt", add_generation_prompt=True)
outputs = model.generate(inputs.to(model.device), max_new_tokens=1536)
print(tokenizer.decode(outputs[0][inputs.shape[1]:], skip_special_tokens=True))

Reproduction

Full code and experiment data: muse-glimmer abliteration repo

Prior work in the same program: gemma-4 abliteration repo

python scripts/capture.py --harmful data/harmful_train.txt \
  --harmless data/harmless_train.txt --out acts/derive.npz --position prompt_final

python scripts/derive_direction.py --acts acts/derive.npz \
  --acts-holdout acts/holdout.npz --out directions/v1

python scripts/abliterate.py --directions directions/v1.npy \
  --out models/mg-abl-s1.0 --scale 1.0

Credits

The method is not original to this work — refusal directions (Arditi et al., 2024), norm-preserving biprojection (grimjim), heretic (p-e-w) and the mlabonne / JailbreakBench / AdvBench corpora. Full attribution in the repo README.

Quantized

TrevorJS/Muse-Glimmer-30B-uncensored-GGUF

Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.