library_name: transformers
license: apache-2.0
pipeline_tag: text-generation
base_model:
- yuxinlu1/gemma-4-12B-coder-fable5-composer2.5-v1
tags: - abliterated
- uncensored
- gemma4
- coding
- code
- reasoning
- thinking
- safetensors
- transformers
Huihui-gemma-4-12B-coder-fable5-composer2.5-v1-abliterated — MLX 4.4 BPW
Mixed-precision MLX quantization of huihui-ai/Huihui-gemma-4-12B-coder-fable5-composer2.5-v1-abliterated, quantized with MLX Smart Quantize (MSQ) — my own sensitivity-based mixed-precision quantization method for Apple Silicon. It measures per-layer NMSE and assigns optimal bit widths automatically, combining architecture knowledge with measured data.
Details
- Type: Vision (VLM)
- Average: 4.45 bits per weight
- Method: MLX Smart Quantize (MSQ)
- AWQ scaling: applied to 96 groups