For privacy reasons a browser tells us at most
"≥ 8 GB RAM, 8 cores" - same reading whether you have 8 GB
or 128 GB. It has no idea how much RAM is free right now, which
apps are open, or whether you have a GPU.
The Abliteration app is integrated with your machine
It reads your exact RAM, GPU model and VRAM,
free memory right now, and picks the sharpest quant that
still fits. Every model page lights up precisely for your rig.
And you can chat with any model, right now
The app is a full local runtime - no API keys, no subscription,
everything runs on your machine. Click any model on this site and
start a conversation in seconds.
No classification signals present. This may not be an abliterated model at all - it could be a repackaging, a merge with unrelated goals, or unrelated content that mentions the term.
no classification signals present (no abliterated, uncensored, or known producer/method markers)
Refusal direction extraction
No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.
apex-flash-1-abliterated is an experimental derivative of apex-flash-1, Cantina Security's open-weights security model developed in partnership with Yeta (@yetalabs on X).
It is intended for authorized security research in environments the researcher owns or has permission to test. The variant modifies refusal behavior broadly; the changes are not limited to security tasks.
Evaluation
This variant has not undergone a separate full-suite evaluation. The results reported for apex-flash-1 apply to the standard checkpoint and should not be attributed to this derivative. Changes in refusal behavior do not establish improvements in security-task performance or reliability.
Capabilities
The checkpoint retains the standard model's image-text-to-text architecture. Image and video performance have not been evaluated for this variant.
Model details
The checkpoint uses the GLM-5.3-Flash architecture and is distributed in BF16. See the standard model card for the training overview and intended worker-model use.
Explore Apex to learn more about Cantina's security research system.
Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.