For privacy reasons a browser tells us at most
"≥ 8 GB RAM, 8 cores" - same reading whether you have 8 GB
or 128 GB. It has no idea how much RAM is free right now, which
apps are open, or whether you have a GPU.
The Abliteration app is integrated with your machine
It reads your exact RAM, GPU model and VRAM,
free memory right now, and picks the sharpest quant that
still fits. Every model page lights up precisely for your rig.
And you can chat with any model, right now
The app is a full local runtime - no API keys, no subscription,
everything runs on your machine. Click any model on this site and
start a conversation in seconds.
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
'abliterated' in name/tags
is_gguf=0 (base model)
no specific method indicators - defaulting to M1 (most common)
Original model: ./models/Qwen3.5-397B-A17B-6bit (local path)
What is Abliteration?
Abliteration is a mechanistic interpretability technique that identifies and orthogonalizes the "refusal direction" in a model's activation space, surgically removing refusal behavior without full fine-tuning.
Abliteration Parameters
Parameter
Value
Base model
./models/Qwen3.5-397B-A17B-6bit
Ablation method
projection
Refusal vector policy
per-layer
Refusal direction method
projected
Ablation strength
2.0
Probed layers
all
PCA components (ablate-k)
1
Attention only
True
Timestamp
2026-03-03T05:02:07 UTC
This model uses a Mixture-of-Experts (MoE) architecture. Abliteration targets only the attention projection weights (q/k/v/o_proj) to preserve expert routing quality.
⚠️ Disclaimer
This model is intended for research, experimentation, and testing purposes only.
This model may produce harmful, offensive, inappropriate, or otherwise objectionable content.
The abliteration process removes safety guardrails that were intentionally built into the original model.
Do not use this model in production systems, consumer-facing applications, or any context where harmful outputs could cause real-world harm.
The authors and contributors of this toolkit bear no responsibility for any misuse of this model or any harm caused by outputs generated by this model.
By using this model, you agree that you are solely responsible for ensuring its use complies with all applicable laws and ethical guidelines.
This model is shared purely for academic and technical exploration of model internals.
README history
1 version
The author's README evolved over time. Click a version to see its content at that point.
2026-03-03Add files using upload-large-folder toolce561952 KB
Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.