← back to catalog · registered 2026-08-22 13:56

AmixDigital/Laguna-S-2.1-Uncensored-oQ4e

AmixDigital 118B MoE second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/AmixDigital%2FLaguna-S-2.1-Uncensored-oQ4e"
Response includes
  • classification m-uncensored
  • files 27
  • hub_downloads_all_time 1,569
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
2K
284 last 30d - stable
Likes
2
Model age
2mo ago
created 2026-08-03
Downloads over time
Now1.6K→from510↑222%
4538871.3K1.8K510 on Aug 51.6K on Oct 11AugSepOct
Aug 5 → Oct 11 · 51 snapshots · spans 67 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

Tags
mlx safetensors laguna oq quantized moe uncensored apple-silicon macos text-generation conversational custom_code

Related

Total size
63.0 GB
Files
27
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-16 10:29

Files by quantization

Auxiliary files 27 files 63.0 GB
model-00008-of-00013.safetensors 5.06 GB 9c534f83 download
model-00009-of-00013.safetensors 5.06 GB a588063f download
model-00012-of-00013.safetensors 5.06 GB 3c8fa840 download
model-00004-of-00013.safetensors 5.06 GB 79f7636b download
model-00006-of-00013.safetensors 5.06 GB 524f9f38 download
model-00007-of-00013.safetensors 5.06 GB 15368576 download
model-00011-of-00013.safetensors 5.06 GB 39daf97f download
model-00005-of-00013.safetensors 5.06 GB 24174b46 download
model-00010-of-00013.safetensors 5.06 GB dd848e05 download
model-00002-of-00013.safetensors 5.06 GB c56f4756 download
model-00003-of-00013.safetensors 5.06 GB f6a33304 download
model-00001-of-00013.safetensors 4.80 GB 069c4415 download
model-00013-of-00013.safetensors 2.53 GB 4f76d583 download
tokenizer.json 6.95 MB 08593aa6 download
model.safetensors.index.json 159 KB f0f2666c download
config.json 92.8 KB ae51b40e download
README.md 57.3 KB 5cfa8763 download
modeling_laguna.py 40.0 KB 9bcc79b4 download
oq_imatrix_report.json 29.8 KB 31dbb0d4 download
configuration_laguna.py 12.8 KB f10d7285 download
tokenizer_config.json 12.7 KB f0c2308a download
LICENSE-APACHE-2.0 11.1 KB 57bc88a1 download
chat_template.jinja 3.87 KB acf45eb4 download
LICENSE 2.56 KB ec297ac5 download
.gitattributes 1.48 KB a6344aac download
generation_config.json 479 B 7421facb download
special_tokens_map.json 214 B 60bc5592 download

README current version from Hugging Face


library_name: mlx
license: openmdw-1.1
license_link: https://huggingface.co/AmixDigital/Laguna-S-2.1-Uncensored-oQ4e/blob/main/LICENSE
pipeline_tag: text-generation
base_model: SC117/Laguna-S-2.1-Uncensored
base_model_relation: quantized
tags:

  • mlx
  • oq
  • quantized
  • moe
  • laguna
  • uncensored
  • apple-silicon
  • macos

Community quantization · Apple Silicon

Laguna-S-2.1-Uncensored-oQ4e

An independent mixed-precision MLX edition for local software development and agentic coding on Apple Silicon.

118B / ~8B activeMixture of Experts
oQ4eMixed precision
63.017 GiB13 MLX shards
128 GB testedM4 Max · macOS
01 / 15

About This Release

Laguna-S-2.1-Uncensored-oQ4e is an independent MLX oQ level 4 enhanced quantization of SC117/Laguna-S-2.1-Uncensored. The original ancestor is poolside/Laguna-S-2.1.

The release targets local software-development and agentic-coding workloads on Apple Silicon while retaining the direct source's uncensored behavior and native thinking configuration.

  1. Source behavior: inherited from the SC117 derivative; no new behavior edit was performed by AmixDigital.
  2. Published artifact: 13 MLX Safetensors shards produced through the oMLX graphical interface.
  3. Optional acceleration: DFlash is paired at runtime and is not merged into these weights.
Notice

Important Notices

  • This is a community quantization, not an official release by Poolside or SC117.
  • No independent benchmark suite or long-context evaluation has been completed for this release.
  • Only the documented 128 GB test environment has been validated; minimum memory is unmeasured.
  • No DFlash draft weights are included.

Upstream results are not quantization results. Refusal and divergence measurements reported by SC117 were not rerun here and must not be attributed to these MLX artifacts.

Audience warning

Uncensored Behavior and Safety

The direct source reduces refusal behavior through an upstream Abliterix-based modification. It may generate harmful, explicit, biased, illegal, misleading, or otherwise unsafe content. This release is not suitable for all audiences.

02 / 15

Model Details

ArchitectureLaguna Mixture-of-Experts, text-to-text
Parameters118B total; approximately 8B active per token upstream
Layers48
Experts256 routed + 1 shared, top-10 routing
ContextConfigured up to 1,048,576 tokens; not long-context tested here
QuantizationoQ level 4 enhanced, MLX affine mixed precision
Weights63.017 GiB · 13 shards
StateQuantized target weights; optional draft distributed separately

Repository Files

The table groups related files while covering every functional artifact in this repository.

FilePurposeRequired byNotes
model-00001-of-00013.safetensors … model-00013-of-00013.safetensorsQuantized target weightsMLX runtime13 shards · 63.017 GiB
model.safetensors.index.jsonTensor-to-shard indexMLX loader · Hugging FaceIncludes verified conceptual parameter metadata
config.jsonArchitecture and quantization configurationMLX · oMLXLaguna custom code · affine mixed precision
configuration_laguna.py · modeling_laguna.pyCustom Laguna implementationCustom-code loadersCarry Apache-2.0 notices
tokenizer.json · tokenizer_config.json · special_tokens_map.jsonTokenizer data and settingsPrompt encoding and decodingInherited from the direct source
chat_template.jinjaChat, thinking, and tool-message formattingConversational inferenceNative thinking remains enabled by default
generation_config.jsonGeneration defaults and companion referenceoMLX runtimeDFlash remains external and optional
oq_imatrix_report.jsonMachine-readable calibration evidenceQuantization auditPersonal cache path redacted; limitations retained
LICENSE · LICENSE-APACHE-2.0Redistribution termsRecipients and redistributorsOpenMDW-1.1 materials · Apache-2.0-noticed code
README.md · .gitattributesModel card and Hub storage rulesHugging Face HubEnglish Lumen card · LFS tracking
03 / 15

Provenance and Lineage

  1. Published artifact: this repository's MLX oQ4e quantized target weights, prepared by AmixDigital.
  2. Direct source: SC117/Laguna-S-2.1-Uncensored at revision 0e9c665.
  3. Original ancestor: poolside/Laguna-S-2.1 at revision 00af5a5.

The optional DFlash draft is a runtime companion—not an ancestor, source checkpoint, or component of the published weights.

04 / 15

Transformation Details

Upstream modification. SC117 describes the direct source as an Abliterix Trial 16 behavior edit of Poolside's BF16 model, including a LoRA merge, MoE router adjustments, and restoration of the default thinking template. This work was not performed by AmixDigital.

MLX quantization. The published artifacts were created through the graphical interface of oMLX 0.5.4. No CLI quantization command was used.

QuantizationMLX affine mixed precision; 4-bit group-size-64 default with 5-, 6-, and 8-bit overrides
Effective storageApproximately 4.605 bits per conceptual parameter, calculated from 67,664,427,590 weight bytes and 117,561,977,600 parameters
Stored dtypesBF16 tensor values and U32 packed quantization data in Safetensors
Precision overrides386 configured tensor overrides: 105 at 5-bit, 1 at 6-bit, and 280 at 8-bit; 245 use group size 64 and 141 use group size 128
Highest-precision tensorslm_head and model.embed_tokens are configured at 8-bit, group size 64
Calibration presetoqe_code_multilingual
Collection1,024 processed samples · sequence length 512
Expert coverage36,084 / 36,096 routes activated · 12 zero-count experts
Report statuscoverage_sufficient: false · collection_sufficient: false

Calibration limitation: the imatrix did not activate every expert route. The maintainer reports that oMLX used a quantized proxy because the BF16 source did not fit in available unified memory.

05 / 15

Requirements and Compatibility

macOS26.5.2 · tested
Apple SiliconMacBook Pro, M4 Max, 16-core CPU · tested
Unified memory128 GB · tested and recommended
Observed memoryApproximately 70 GB or less · maintainer observation, not instrumented
oMLX0.5.4 · tested
mlx-lm0.31.3 · installed in test environment
mlx-vlm0.6.3 · installed in test environment
MLX coreExact version not recorded

Minimum unified memory has not been measured. Disk size alone is not a safe minimum-memory estimate.

06 / 15

Usage

Verified oMLX workflow

  1. Add the obtained model directory to oMLX 0.5.4.
  2. Load the model on a compatible Apple Silicon Mac.
  3. Keep thinking enabled when the task benefits from reasoning.
  4. Run a short software-development prompt before increasing context or enabling optional acceleration.

The maintainer's smoke test requested a landing-page implementation. The model loaded successfully, produced a coherent result, and exposed thinking behavior.

Prompting, Thinking and Tool Use

The chat template defaults enable_thinking to true. Thinking was observed, but tool calling was not tested on this quantization.

07 / 15

Acceleration and Companion Models

This target can be paired with the separately distributed poolside/Laguna-S-2.1-DFlash draft. DFlash proposes token blocks that the target verifies; it does not replace or modify the target weights.

Open the DFlash companion repository →

  1. Obtain the official draft repository separately.
  2. Enable DFlash in the target's oMLX settings.
  3. Select the draft directory and reload the target.

Runtime pairing was completed in oMLX 0.5.4. Throughput, acceptance rate, equivalence, and long-context behavior were not measured; no acceleration factor is claimed.

08 / 15

Evaluation and Performance

No independent benchmark suite has been completed for this release.

CheckResultLimit
Model loadSuccessfulOne Mac configuration
GenerationCoherent landing pageSingle qualitative prompt
ThinkingObservedNo structured evaluation
Tool callingNot testedInherited configuration only
Long contextNot testedConfigured maximum unvalidated
DFlashRuntime pairing completedSpeed and acceptance unmeasured

Benchmarks from Poolside or SC117 describe their own artifacts and must not be interpreted as measurements of this quantization.

09 / 15

Intended Uses

  • Local software development and code-oriented assistance on compatible Apple Silicon Macs.
  • Agentic-coding experiments with human review and constrained tool permissions.
  • General text generation where unvalidated general-purpose quality is understood.
10 / 15

Out-of-Scope Uses

  • Unsupervised high-stakes medical, legal, financial, employment, safety, or infrastructure decisions.
  • Autonomous destructive tool execution without sandboxing, review, and recovery controls.
  • Uses that violate law, rights, license terms, or platform policies.
  • Claims of Windows, Linux, CUDA, GGUF, vLLM, or server-GPU compatibility.
11 / 15

Risks, Biases and Limitations

  • Uncensored output may be harmful, explicit, biased, fabricated, or unsafe.
  • Quantization may add variance or degradation beyond upstream limitations.
  • Twelve of 36,096 expert routes were not activated during calibration.
  • Quality, tool use, long context, minimum memory, and general stability remain unmeasured.
  • Custom model code requires security review before remote-code execution.
12 / 15

Recommendations and Mitigations

  • Evaluate the exact model and prompts before deployment.
  • Keep human review for code changes, tool use, and consequential outputs.
  • Sandbox generated code and restrict credentials and filesystem access.
  • Apply access controls and content filtering where necessary.
  • Start with short contexts and monitor memory pressure on untested hardware.
13 / 15

License and Responsible Use

The published Model Materials are distributed under the OpenMDW-1.1 License. The complete inherited agreement is included as LICENSE. Redistribution must retain the agreement and applicable copyright and origin notices.

The redistributed Python implementation files also carry Apache-2.0 notices. A complete copy is included as LICENSE-APACHE-2.0.

Users remain responsible for third-party rights, applicable law, and determining whether the model and its outputs are appropriate for a particular use. AmixDigital and the maintainers are not responsible for illegal, abusive, or otherwise unauthorized uses of the model or its outputs.

14 / 15

Acknowledgements

  • Poolside for the original Laguna S 2.1 and official DFlash companion.
  • SC117 for the direct uncensored source and behavior-edit provenance.
  • Abliterix contributors for the upstream behavior-edit method.
  • oMLX contributors for MLX serving and oQ quantization tooling.
15 / 15

Model Card Contact

Contact AmixDigital on Hugging Face or open a discussion in this model repository to report an error or propose a correction.

This is an independent community quantization. It is not an official release by, affiliated with, or endorsed by Poolside or SC117. All upstream models, methods, names, and trademarks remain the property of their respective owners.

README history 12 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-16Fix Laguna model card dark theme4bc539157.3 KB
    Loading...
  2. 2026-08-16Apply official Laguna visual direction to model cardd9a131e57.3 KB
    Loading...
  3. 2026-08-04Adapt Lumen model card for dark mode719de5057.2 KB
    Loading...
  4. 2026-08-03Remove audience gate and clarify user responsibilityad3c6e950 KB
    Loading...
  5. 2026-08-03Finalize pre-publication documentation and licensesef3029b49.9 KB
    Loading...
  6. 2026-08-03Add prominent DFlash companion linkd0c30cc41.3 KB
    Loading...
  7. 2026-08-03Unify all model card sections as Lumen panels721f5fa40.8 KB
    Loading...
  8. 2026-08-03Remove rendered HTML inventory and simplify model estimatesf76e65840.2 KB
    Loading...
  9. 2026-08-03Apply Amix Lumen model card style572df1748.8 KB
    Loading...
  10. 2026-08-03Refine Vellum proportions from live Hugging Face QAa8ff96048.9 KB
    Loading...
  11. 2026-08-03Improve Vellum spacing and section hierarchy10eb38049.1 KB
    Loading...
  12. 2026-08-03Add files using upload-large-folder tool354564447.6 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration