← back to catalog · registered 2026-08-22 13:56

ToastyPigeon/g3-12b-it-unalign

ToastyPigeon Gemma 12B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/ToastyPigeon%2Fg3-12b-it-unalign"
Response includes
  • classification unknown
  • files 16
  • benchmarks 5 entries
  • author_summary 6 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
?
Primary method

Unclassified

No clear signals of an abliteration technique in this model.
Confidence
UNKNOWN
Why this label 1 signal
No classification signals present. This may not be an abliterated model at all - it could be a repackaging, a merge with unrelated goals, or unrelated content that mentions the term.
  • no classification signals present (no abliterated, uncensored, or known producer/method markers)
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
12
↑ 443% in 90 days
Likes
0
Model age
18mo ago
created 2025-03-21

Training datasets

Corpora the author lists in the model card. Datasets tracked in our /datasets catalog carry a category badge linking to the workflow stage. Others open on Hugging Face.

Downloads over time
Now76→from14↑443%
028568314 on Mar 19, 202576 on Oct 1176 on Oct 7Mar '25Jun '25Sep '25Dec '25MarJunSep
Mar 19, 2025 → Oct 11 · 121 snapshots · spans 571 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
Arena-Battles 3976 LM-Arena
LM Arena Elo 1335.3304642871612 LM-Arena
Arena-Elo-Lower 1326.1060034720686 LM-Arena
Arena-Elo-Upper 1344.5549251022537 LM-Arena
Arena-Rank 49 LM-Arena

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
gemma
Tags
transformers safetensors gemma3_text text-generation axolotl generated_from_trainer conversational dataset:ToastyPigeon/unalign-v2 base_model:unsloth/gemma-3-12b-it base_model:finetune:unsloth/gemma-3-12b-it license:gemma text-generation-inference

Related

Total size
21.9 GB
Files
16
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2025-03-22 18:46

Files by quantization

Auxiliary files 16 files 22.0 GB
model-00003-of-00005.safetensors 4.59 GB 8600858c download
model-00004-of-00005.safetensors 4.59 GB 93cde622 download
model-00002-of-00005.safetensors 4.59 GB c6b6d533 download
model-00001-of-00005.safetensors 4.58 GB d81136bc download
model-00005-of-00005.safetensors 3.56 GB cf883d12 download
training_args.bin 6.43 KB 6d558d9b download
tokenizer.json 31.8 MB 4667f208 download
tokenizer.model 4.47 MB 1299c11d download
tokenizer_config.json 1.10 MB c3982b0b download
model.safetensors.index.json 51.4 KB d4e4efa6 download
README.md 2.29 KB 19fd2907 download
.gitattributes 1.53 KB 52373fe2 download
config.json 937 B 7288216c download
special_tokens_map.json 670 B bdd437b8 download
generation_config.json 213 B 0e61dc62 download
added_tokens.json 35.0 B e17bde03 download

README current version from Hugging Face


library_name: transformers
license: gemma
base_model: unsloth/gemma-3-12b-it
tags:

  • axolotl
  • generated_from_trainer
    datasets:
  • ToastyPigeon/unalign-v2
    model-index:
  • name: g3-12b-it-unalign
    results: []

g3-12b-it-unalign

This model is a fine-tuned version of unsloth/gemma-3-12b-it on the ToastyPigeon/unalign-v2 dataset.
It achieves the following results on the evaluation set:

  • Loss: 1.4684

So, it seems alright. I noticed however that the responses got pretty short at the end of the 2nd epoch. Not like, unusably short, but generally shorter than I personally like.

The epoch 1 test gguf is based on this commit.

I personally prefer epoch 1 to epoch 2, and will likely update this or make a second proper commit for epoch 1.

Update: I did indeed make a second commit for the epoch 1 checkpoint.

Training hyperparameters

The following hyperparameters were used during training:

  • learning_rate: 1e-05
  • train_batch_size: 4
  • eval_batch_size: 4
  • seed: 69
  • optimizer: Use apollo_adamw_layerwise with betas=(0.9,0.999) and epsilon=1e-08 and optimizer_args=proj=random,rank=1,scale=128.0,scale_type=tensor,update_proj_gap=200
  • lr_scheduler_type: cosine
  • lr_scheduler_warmup_steps: 5
  • num_epochs: 2.0

Training results

Training Loss Epoch Step Validation Loss
7.8965 0.0118 1 6.4897
4.4934 0.2 17 4.0497
3.9523 0.4 34 3.7484
3.5624 0.6 51 3.3152
2.7168 0.8 68 2.4773
2.1303 1.0 85 1.9483
1.8215 1.2 102 1.7577
1.7199 1.4 119 1.6561
1.5771 1.6 136 1.5611
1.5599 1.8 153 1.5124
1.4831 2.0 170 1.4684

Framework versions

  • Transformers 4.50.0.dev0
  • Pytorch 2.5.1+cu124
  • Datasets 3.4.1
  • Tokenizers 0.21.1

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2025-03-22Update README.md63d08402.3 KB
    Loading...
  2. 2025-03-22Update README.md5f234942.1 KB
    Loading...
  3. 2025-03-21End of training2910a657 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration