← back to catalog · registered 2026-08-22 13:56

zaakirio/LFM2.5-1.2B-Instruct-Uncensored-GGUF

zaakirio Lfm 1.2B GGUF second-order 128K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/zaakirio%2FLFM2.5-1.2B-Instruct-Uncensored-GGUF"
Response includes
  • classification m8
  • files 7
  • hub_downloads_all_time 7,071
  • author_summary 11 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
7K
749 last 30d - stable
Likes
5
Model age
4mo ago
created 2026-06-01
Downloads over time
Now7.3K→from841↑768%
5183K5.5K7.9K841 on Jun 107.3K on Oct 11JunJulAugSepOct
Jun 10 → Oct 11 · 57 snapshots · spans 123 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Variants by this author 2 formats · 1K downloads combined

The same weights this author released in different packaging. Pick the format that matches your runtime.

Metadata

License
other
Languages
en ar zh ja ko ru
Quantizations
Q3_K Q4_K Q5_K Q6_K Q8_0
Tags
gguf heretic abliterated decensored uncensored liquid lfm2 lfm2.5 edge llama.cpp conversational text-generation

Related

Total size
4.08 GB
Files
7
Quantizations
6
Registered
2026-08-22 13:56
Last updated on HF
2026-06-04 18:05

Files by quantization

Q8_0 1 file 1.16 GB
LFM2.5-1.2B-Instruct-Uncensored-Q8_0.gguf 1.16 GB 7204d64d download
Q6_K 1 file 918 MB
LFM2.5-1.2B-Instruct-Uncensored-Q6_K.gguf 918 MB a446d36d download
Q5_K 1 file 804 MB
LFM2.5-1.2B-Instruct-Uncensored-Q5_K_M.gguf 804 MB 0d583286 download
Q4_K 1 file 697 MB
LFM2.5-1.2B-Instruct-Uncensored-Q4_K_M.gguf 697 MB 68acd89a download
Q3_K 1 file 573 MB
LFM2.5-1.2B-Instruct-Uncensored-Q3_K_M.gguf 573 MB 22eb750a download
Auxiliary files 2 files 6.14 KB
README.md 4.27 KB b85db924 download
.gitattributes 1.87 KB 0f1c8773 download

README current version from Hugging Face


base_model: zaakirio/LFM2.5-1.2B-Instruct-Uncensored
base_model_relation: quantized
quantized_by: zaakirio
license: other
license_name: lfm1.0
license_link: https://huggingface.co/LiquidAI/LFM2.5-1.2B-Instruct/blob/main/LICENSE
library_name: gguf
pipeline_tag: text-generation
language:

  • en
  • ar
  • zh
  • ja
  • ko
  • ru
    tags:
  • heretic
  • abliterated
  • decensored
  • uncensored
  • liquid
  • lfm2
  • lfm2.5
  • edge
  • gguf
  • llama.cpp
  • conversational

LFM2.5-1.2B-Instruct-Uncensored — GGUF

GGUF quantizations of zaakirio/LFM2.5-1.2B-Instruct-Uncensored,
a decensored (Heretic-abliterated) version of
LiquidAI/LFM2.5-1.2B-Instruct.

These files run with llama.cpp and any
tool built on it e.g. LM Studio, Jan, koboldcpp, etc.

Requires a recent llama.cpp build. LFM2 is a hybrid (convolution + attention)
architecture; only llama.cpp builds that include LFM2 support can load these
files. Use a current release (or current LM Studio / Jan). Older builds will
fail with an "unknown architecture 'lfm2'" error.

Files

File Quant Size Notes
LFM2.5-1.2B-Instruct-Uncensored-Q3_K_M.gguf Q3_K_M 573 MB Smallest; lowest quality. For very tight memory.
LFM2.5-1.2B-Instruct-Uncensored-Q4_K_M.gguf Q4_K_M 697 MB Recommended — best size/quality balance.
LFM2.5-1.2B-Instruct-Uncensored-Q5_K_M.gguf Q5_K_M 804 MB Higher quality, slightly larger.
LFM2.5-1.2B-Instruct-Uncensored-Q6_K.gguf Q6_K 918 MB Near-lossless.
LFM2.5-1.2B-Instruct-Uncensored-Q8_0.gguf Q8_0 1.2 GB Effectively lossless vs the BF16 source.

Not sure which to pick? Start with Q4_K_M. Go up to Q5/Q6/Q8 if you have the
memory and want maximum fidelity; drop to Q3 only if you're memory-constrained.

Usage

llama.cpp (auto-download from this repo)

# Interactive chat — downloads the chosen quant automatically
llama-cli -hf zaakirio/LFM2.5-1.2B-Instruct-Uncensored-GGUF:Q4_K_M

# OpenAI-compatible server
llama-server -hf zaakirio/LFM2.5-1.2B-Instruct-Uncensored-GGUF:Q4_K_M -c 4096

Or, with a file you've already downloaded:

llama-cli -m LFM2.5-1.2B-Instruct-Uncensored-Q4_K_M.gguf -p "Hello, who are you?"

LM Studio / Jan

Search for zaakirio/LFM2.5-1.2B-Instruct-Uncensored-GGUF in the in-app model
browser, or download a .gguf file from this page and load it.

Download a single file

pip install -U "huggingface_hub[cli]"
hf download zaakirio/LFM2.5-1.2B-Instruct-Uncensored-GGUF \
  --include "LFM2.5-1.2B-Instruct-Uncensored-Q4_K_M.gguf" --local-dir ./

Prompt format

The chat template is embedded in the GGUF files, so chat-aware tools apply it
automatically. For reference, it is ChatML-style:

<|startoftext|><|im_start|>system
{system_prompt}<|im_end|>
<|im_start|>user
{prompt}<|im_end|>
<|im_start|>assistant

About the base model

This is a decensored derivative produced with Heretic
(automatic directional ablation). Compared with the original LFM2.5-1.2B-Instruct:

Metric Decensored Original
Refusals (/100 harmful prompts) 5 98
KL divergence (harmless prompts) 0.1003 0 (by definition)

See the source model card
for the full abliteration parameters and run details.

Intended use & disclaimer

This model has had its refusal behavior substantially removed and will comply
with requests the original model would have declined. It is provided for
research and unrestricted local use. You are responsible for how you use it
and for complying with all applicable laws and with the base model's
lfm1.0 license,
which carries over to this derivative.

Provenance

  • Quantized from zaakirio/LFM2.5-1.2B-Instruct-Uncensored (BF16) using llama.cpp convert_hf_to_gguf.py + llama-quantize.
  • Base model: LiquidAI/LFM2.5-1.2B-Instruct
  • Decensoring tool: Heretic by p-e-w

README history 5 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-06-04Upload README.md with huggingface_hub95240134.3 KB
    Loading...
  2. 2026-06-01Update README.md3bd27284.4 KB
    Loading...
  3. 2026-06-01Update README.md60e80c34.4 KB
    Loading...
  4. 2026-06-01Improve GGUF card: lfm2.5/liquid/edge tags, -hf usage, prompt format, llama.c...76f17bb4.4 KB
    Loading...
  5. 2026-06-01Add Q3_K_M/Q4_K_M/Q5_K_M/Q6_K/Q8_0 GGUF quants of LFM2.5-1.2B-Instruct-Uncens...d85c6573.4 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration