← back to catalog · registered 2026-08-22 13:56

Silicone-Moss/MistralAI-Magistral-Small-2507-Heretic-Uncensored

Silicone-Moss Mistral 24B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Silicone-Moss%2FMistralAI-Magistral-Small-2507-Heretic-Uncensored"
Response includes
  • classification m3
  • files 8
  • hub_downloads_all_time 93
  • author_summary 4 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M3
Primary method

Layer-wise ablation

Applied on top of direct removal inherited from the base model.
Confidence
HIGH
Inherited from base model
Why this label 2 signals
Producer identity confirmed by naming conventions, tags or the model card. This label is very unlikely to change.
  • 'heretic' in model name (Heretic-produced)
  • Heretic uses layer-wise optimization (M3) with underlying direction removal (M1)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
HIGH
Why we say so
name contains 'heretic'; Heretic default extraction is difference-of-means (Arditi 2024)
Downloads · lifetime
93
33 last 30d - stable
Likes
1
Descendants
2
in 2 direct forks
Model age
8mo ago
created 2026-02-11
Downloads over time
Now103→from26↑296%
22528111126 on Feb 18103 on Oct 11FebAprJunAugOct
Feb 18 → Oct 11 · 73 snapshots · spans 235 days

Genealogy 2 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en
Tags
safetensors mistral ablation research nlp 24b heretic magistral roleplay text-generation conversational en

Related

Total size
43.9 GB
Files
8
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-02-16 21:17

Files by quantization

Auxiliary files 8 files 43.9 GB
model.safetensors 43.9 GB 68450e53 download
tokenizer.json 16.3 MB 5bc988cc download
README.md 3.21 KB 8250d521 download
chat_template.jinja 1.56 KB c59aca5c download
.gitattributes 1.53 KB 52373fe2 download
config.json 693 B 980f27aa download
tokenizer_config.json 338 B db9fb605 download
generation_config.json 110 B 1103a7d5 download

README current version from Hugging Face


license: apache-2.0
language:

  • en
    pipeline_tag: text-generation
    tags:
  • ablation
  • research
  • nlp
  • 24b
  • mistral
  • heretic
  • magistral
  • roleplay
    base_model:
  • mistralai/Magistral-Small-2507

MistralAI-Magistral-Small-2507-Heretic

[!CAUTION]
EXPERIMENTAL RESEARCH ARTIFACT

This model represents an aggressive application of the Heretic repository and optimization methodology.

  • Status: STILL TESTING / BETA
  • Behavior: This model has significantly reduced refusal mechanisms. It recorded only 6 refusals (out of 100) in the test set.
  • Use Case: This is a research artifact intended for testing the limits of vector-based intervention. Use with appropriate caution.

Model Summary

MistralAI-Magistral-Small-2507-Heretic is a fine-tuned language model resulting from the Heretic repository and optimization methodology. It utilizes a targeted vector intervention technique (orthogonalization/abliteration) tuned via Optuna to minimize refusal responses while maintaining high coherence.

This specific checkpoint represents Trial 116, which achieved a low refusal count with a KL Divergence of ~0.0124. This indicates exceptional adherence to the base model's probability distribution.

Run Configuration: "Trial 116"

The following parameters define the intervention vector applied to the model. This configuration was discovered during the hyperparameter search.

Optimization Results

Metric Value Description
Refusal Count 6 The model refused 6 prompts in the Heretic test set (approx. 6% refusal rate).
KL Divergence 0.0124 Measures deviation from the base model's probability distribution. (Lower is better).
Trial ID 115 Specific Optuna trial identifier.
Direction Scope Global The refusal vector was calculated once globally and applied across layers.

Intervention Parameters

Interventions were applied to two primary distinct layers: the Attention Output Projection (attn.o_proj) and the MLP Down Projection (mlp.down_proj).

Parameter Scope Setting Value
Attention Output attn.o_proj.max_weight 1.495
(attn.o_proj) attn.o_proj.max_weight_position 26.75 (Layer Depth)
attn.o_proj.min_weight 1.393
attn.o_proj.min_weight_distance 24.81
MLP Down Proj mlp.down_proj.max_weight 1.148
(mlp.down_proj) mlp.down_proj.max_weight_position 33.08 (Layer Depth)
mlp.down_proj.min_weight 0.319
mlp.down_proj.min_weight_distance 10.28

Usage & Limitations

  • Intended Use: Research into model alignment, vector arithmetic, and uninhibited creative writing.
  • Risks: This base model has removed most safety guardrails removed. It may generate content for sensitive prompts that the base model would refuse. Thank you for trying my experiments.

Credits & References

This research builds upon the excellent work of the open-source AI community:

README history 5 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-02-16Update README.md6fefba43.2 KB
    Loading...
  2. 2026-02-16Update README.mdfb562a53.2 KB
    Loading...
  3. 2026-02-16Update README.md9cf65363.2 KB
    Loading...
  4. 2026-02-16Update README.md912d6a33.3 KB
    Loading...
  5. 2026-02-16Create README.md7861dab3.4 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration