license: apache-2.0
base_model: microsoft/Phi-3-mini-4k-instruct
tags:
- uncensored
- abliterated
- gguf
- phi
- conversational
pipeline_tag: text-generation
phi-3-mini-uncensored
Uncensored variant of microsoft/Phi-3-mini-4k-instruct.
Method
- Abliteration (strength=0.2) — refusal direction removed from all layers
- LoRA fine-tune on
Guilherme34/uncensor(2 epochs, r=16, alpha=32) - Re-abliteration (strength=0.35) — stronger pass to remove residual refusals
Eval Results
| Split | Refused |
|---|---|
| Harmful (64 prompts) | 6/64 |
| Harmless (64 prompts) | 0/64 |
Usage
llama-cli -m phi_3_mini_uncensored.Q4_K_M.gguf -p "Your prompt here"
Training Config
| Parameter | Value |
|---|---|
| Base Model | microsoft/Phi-3-mini-4k-instruct |
| Fine-tune Dataset | Guilherme34/uncensor |
| Epochs | 2 |
| LoRA r | 16 |
| LoRA alpha | 32 |
| Learning Rate | 0.0002 |
| Abliteration Strength | 0.2 |
| Re-abliteration Strength | 0.35 |
Credits
- Abliteration technique: andyrdt/refusal_direction
- Weight editing: Sumandora/remove-refusals-with-transformers