← back to catalog · registered 2026-08-22 13:56

Rootkit7/Laguna-S-2.1-uncensored-GGUF

Rootkit7 GGUF MoE second-order 1.0M ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Rootkit7%2FLaguna-S-2.1-uncensored-GGUF"
Response includes
  • classification m-uncensored
  • files 6
  • hub_downloads_all_time 91
  • author_summary 11 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
91
10 last 30d - stable
Likes
0
Model age
2mo ago
created 2026-08-07
Downloads over time
Now91→from15↑507%
1140699915 on Aug 591 on Oct 1191 on Sep 25AugSepOct
Aug 5 → Oct 11 · 50 snapshots · spans 67 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Variants by this author 2 formats · 10 downloads combined

The same weights this author released in different packaging. Pick the format that matches your runtime.

Metadata

Quantizations
Q4_K Q5_K Q6_K Q8_0
Tags
gguf abliteration uncensored moe llama.cpp base_model:Rootkit7/Laguna-S-2.1-uncensored base_model:quantized:Rootkit7/Laguna-S-2.1-uncensored license:openmdw-1.1 endpoints_compatible region:us conversational

Related

Total size
350 GB
Files
6
Quantizations
5
Registered
2026-08-22 13:56
Last updated on HF
2026-08-07 22:42

Files by quantization

Q8_0 1 file 116 GB
Laguna-S-2.1-uncensored-Q8_0.gguf 116 GB ******** download
Q6_K 1 file 89.9 GB
Laguna-S-2.1-uncensored-Q6_K.gguf 89.9 GB ******** download
Q5_K 1 file 77.7 GB
Laguna-S-2.1-uncensored-Q5_K_M.gguf 77.7 GB ******** download
Q4_K 1 file 66.3 GB
Laguna-S-2.1-uncensored-Q4_K_M.gguf 66.3 GB ******** download
Auxiliary files 2 files 5.33 KB
README.md 3.44 KB 0721b4d9 download
.gitattributes 1.89 KB b346a7fa download

README current version from Hugging Face


base_model: Rootkit7/Laguna-S-2.1-uncensored
license: openmdw-1.1
tags:

  • abliteration
  • uncensored
  • moe
  • gguf
  • llama.cpp

Laguna-S-2.1-uncensored — GGUF

GGUF quantizations of Rootkit7/Laguna-S-2.1-uncensored
— an uncensored (safety-refusal-removed) build of poolside's Laguna-S-2.1, a Mixture-of-Experts
reasoning model. The refusals are removed while capability is preserved: on a held-out multilingual harmful
set, refusal drops from ~95% → 3% and Solutus's automatic capability gate passes (no measurable
degradation). Built with the Solutus abliteration toolkit — the exact recipe is under Provenance & safety below.

⚠️ Runtime requirement — read this first. Laguna's hybrid architecture needs llama.cpp
with Laguna support
: poolside's fork (git clone --branch laguna https://github.com/poolsideai/llama.cpp) or a recent-enough upstream (support is in
ggml-org/llama.cpp #25165, ~release b10087). Older llama.cpp — and current Ollama /
LM Studio — will not load these files yet.

ℹ️ Measured & gate-verified. Laguna is now a whitelisted architecture in Solutus; these numbers passed
the honest capability gate and every quant here was run-validated (see below). They are measured on
Laguna-S-2.1 specifically — indicative for that model, not a cross-model certification.

✅ Validation status — all quants run-validated. Q4_K_M, Q5_K_M, Q6_K, and Q8_0 were each loaded on the
poolside llama.cpp@laguna fork and checked live: they load and serve, comply on a harmful prompt
(abliteration carried through), stay coherent on a benign prompt, and reasoning works (a separate
reasoning_content thinking trace plus the correct final answer). For full precision, use the safetensors
repo (linked below).

Quants

File Quant Size Notes
Laguna-S-2.1-uncensored-Q4_K_M.gguf Q4_K_M ~71 GB best size/quality balance (recommended)
Laguna-S-2.1-uncensored-Q5_K_M.gguf Q5_K_M ~83 GB higher quality
Laguna-S-2.1-uncensored-Q6_K.gguf Q6_K ~97 GB near-lossless
Laguna-S-2.1-uncensored-Q8_0.gguf Q8_0 ~125 GB highest-fidelity quant

All quants (Q4_K_M / Q5_K_M / Q6_K / Q8_0) are run-validated — see the validation note above. For
full precision, use the safetensors model at
Rootkit7/Laguna-S-2.1-uncensored.

Run

git clone --branch laguna https://github.com/poolsideai/llama.cpp
cd llama.cpp && cmake -B build -DGGML_CUDA=ON && cmake --build build -j --target llama-cli

# chat (reasoning on by default — this is a thinking model; use sampling, not greedy):
./build/bin/llama-cli -m Laguna-S-2.1-uncensored-Q4_K_M.gguf --jinja --temp 0.7
# reasoning off for a direct answer:
./build/bin/llama-cli -m Laguna-S-2.1-uncensored-Q4_K_M.gguf --jinja -rea off --temp 0.7 -p "..."

Provenance & safety

  • Base: poolside/Laguna-S-2.1 @ 00af5a51782109b587a3b3bbf11875e566036fa7; recipe: Solutus ega,
    union of 8 datasets, α=5, router_scale=0.74.
  • Safety refusals removed — a research artifact. It will comply with harmful requests. Use responsibly
    and per the base model's license, OpenMDW-1.1 (inherited from poolside/Laguna-S-2.1; a permissive
    open-weights license allowing use, modification, and redistribution incl. derivatives).
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration