← back to catalog · registered 2026-08-22 13:56

GrEarl/Kimi-K3-Abliterated-V1-Q2_K-GGUF

GrEarl Kimi GGUF MoE second-order 1.0M ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/GrEarl%2FKimi-K3-Abliterated-V1-Q2_K-GGUF"
Response includes
  • classification m8
  • files 98
  • hub_downloads_all_time 216
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
216
99 last 30d - stable
Likes
11
Model age
2mo ago
created 2026-08-04
Downloads over time
Now219→from27↑711%
179116523827 on Aug 5219 on Oct 11219 on Oct 4AugSepOct
Aug 5 → Oct 11 · 50 snapshots · spans 67 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
other
Languages
en ja zh
Tags
gguf kimi-k3 q2_k q4_k moe abliterated uncensored text-generation en ja zh base_model:Uniboshi/Kimi-K3-Abliterated-V1

Related

Total size
865 GB
Files
98
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-04 21:57

Files by quantization

Auxiliary files 98 files 865 GB
Kimi-K3-Q2_K-00011-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00013-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00014-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00015-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00017-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00018-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00019-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00021-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00022-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00023-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00025-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00026-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00027-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00029-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00030-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00031-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00033-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00034-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00035-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00037-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00038-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00039-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00041-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00042-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00043-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00045-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00046-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00047-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00049-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00050-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00051-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00053-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00054-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00055-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00057-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00058-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00059-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00061-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00062-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00063-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00065-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00066-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00067-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00069-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00070-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00071-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00073-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00074-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00075-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00077-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00078-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00079-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00081-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00082-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00083-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00085-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00086-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00087-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00089-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00090-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00091-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00002-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00003-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00005-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00006-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00007-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00009-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00010-of-00094.gguf 9.40 GB ******** download
Kimi-K3-Q2_K-00012-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00016-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00020-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00024-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00028-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00032-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00036-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00040-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00044-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00048-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00052-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00056-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00060-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00064-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00068-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00072-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00076-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00080-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00084-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00088-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00092-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00093-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00004-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00008-of-00094.gguf 9.30 GB ******** download
Kimi-K3-Q2_K-00094-of-00094.gguf 1.78 GB ******** download
Kimi-K3-Q2_K-00001-of-00094.gguf 637 MB ******** download
release-manifest.json 26.1 KB 3eaa33c5 download
README.md 8.58 KB a4e7e884 download
.gitattributes 7.82 KB 5d8f9618 download
LICENSE 2.99 KB 0de9f069 download

README current version from Hugging Face


license: other
license_name: kimi-k3
base_model: Uniboshi/Kimi-K3-Abliterated-V1
base_model_relation: quantized
pipeline_tag: text-generation
tags:

  • gguf
  • kimi-k3
  • q2_k
  • q4_k
  • moe
  • abliterated
  • uncensored
    language:
  • en
  • ja
  • zh
    extra_gated_prompt: |
    This is a quantized derivative of an intentionally reduced-refusal model.
    Before requesting access, you must read the model cards for
    Uniboshi/Kimi-K3-Abliterated-V1, moonshotai/Kimi-K3, and
    GrEarl/Kimi-K3-GGUF, together with the complete Kimi K3 License. Request
    access only if you understand the reduced-safeguard nature of the model and
    accept responsibility for lawful, controlled use.
    extra_gated_fields:
    "I have read all three required upstream and quantization model cards": checkbox
    "I have read and accept the Kimi K3 License and applicable upstream conditions": checkbox
    "I understand that this model is not a safety boundary and may produce harmful content": checkbox
    "I accept responsibility for access control, output review, deployment, and legal compliance": checkbox

Kimi-K3-Abliterated-V1 Q2_K GGUF

A 94-part mixed Q2_K/Q4_K GGUF quantization of
Uniboshi/Kimi-K3-Abliterated-V1,
built with the quantization policy and split layout of
GrEarl/Kimi-K3-GGUF v3.

[!WARNING]
The base model intentionally reduces refusal behavior. It may generate
inaccurate, offensive, unsafe, or illegal suggestions. This model is not a
safety boundary. Users are responsible for access controls, output review,
downstream use, and compliance with applicable law and license terms.

Required reading and acknowledgement

Do not request access, download, run, redistribute, or deploy this model until
you have read these resources in full:

  1. Uniboshi/Kimi-K3-Abliterated-V1 —
    the full-weight base model, its reduced-refusal intent, usage notes, and
    limitations.
  2. moonshotai/Kimi-K3 — the
    original architecture, intended deployment guidance, and
    Kimi K3 License.
  3. GrEarl/Kimi-K3-GGUF — the Q2
    v3 quantization design, llama.cpp runtime requirements, measured evidence,
    and quantization limitations reused by this build.

By requesting access or using this repository, you confirm that you have read
those cards and the complete license, understand that safeguards have
intentionally been weakened, and accept responsibility for controlled and
lawful use. This acknowledgement does not replace or modify the upstream
license. Under the Hugging Face gated-model workflow, an access request also
shares the requester's account identity and contact information with the
repository owner.

Lineage and release identity

Item Value
Hugging Face base model Uniboshi/Kimi-K3-Abliterated-V1
HF relationship quantized
Original model moonshotai/Kimi-K3
Structural and quantizer template GrEarl/Kimi-K3-GGUF v3
Parts 94
Total size 864.81 GiB
Quantization layout mixed Q2_K/Q4_K
Architecture kimi-k3
Modality text only

This release is a quantized derivative of Uniboshi's checkpoint.
It is not a fine-tune of GrEarl/Kimi-K3-GGUF.
The release build does not apply a refusal vector.
It does not apply a QAware correction or reconstruct any weight from a rank-one
approximation. No experimental post-quantization abliteration contributes to
the published bytes.

Direct quantization method

The build uses a change-aware materialization strategy. The pinned Uniboshi
checkpoint changes 280 writer tensors relative to its Kimi-K3 source lineage.
One changed tensor is the vision-only mm_projector.proj.2.weight, which is
outside this text GGUF. The remaining 279 text tensors are fetched directly
from the pinned Uniboshi BF16 safetensors and quantized exactly once from BF16
to Q4_K.

Uniboshi tensor family GGUF role Count Output type
embed_tokens token_embd.weight 1 Q4_K
o_proj attention output writers 93 Q4_K
down_proj dense/shared MLP down writers 93 Q4_K
routed_expert_up_proj Stable LatentMoE routed up writers 92 Q4_K
Total 279 Q4_K

The exact BF16 input for those tensors is 29.8457 GiB and their packed Q4_K
payload is 8.3941 GiB. Tensors established as unchanged in the pinned
full-weight comparison reuse the byte-identical v3 quantization payload for the
same source weight. This avoids downloading and re-quantizing several
terabytes of unchanged data while producing the intended Uniboshi quantized
checkpoint. It is a build optimization, not a model merge or a transfer of an
inferred refusal direction.

The v3 policy keeps quality-sensitive writers in Q4_K and the large routed
expert stacks in Q2_K. None of the Q2_K routed-expert payloads is modified by
the 279-tensor direct-quantization pass.

Integrity verification

The release pipeline fails closed on revision drift, tensor-name or shape
mismatch, unexpected quantization type, packed-size mismatch, and incomplete
downloads. For every output part it records and checks:

  • the pinned source LFS size and SHA-256;
  • the exact Uniboshi BF16 byte range and SHA-256 for every materialized tensor;
  • the output payload SHA-256 for every directly quantized tensor;
  • unchanged payload hashes against the structural template;
  • file size, tensor metadata, split metadata, and a GGUFReader reopen;
  • the final full-file SHA-256.
Audited item Result
Pinned GGUF template parts 94 / 94
Output parts reopened with GGUFReader 94 / 94
Directly quantized Uniboshi text tensors 279
Q2_K routed-expert payloads modified 0
Output file hashes recorded 94 / 94

The repository includes release-manifest.json with all 94 file sizes,
output hashes, template-source hashes, pinned revisions, and the materialization
declaration.

Runtime and evaluation status

Runtime, throughput, refusal behavior, and quality results from the retired
experimental QAware/refusal-transfer candidate do not apply to this release
and are not claimed here. At initial publication, this exact uploaded revision
has passed the weight-space and file-integrity gates above but has not completed
an end-to-end runtime or behavioral evaluation. Runtime timings, raw outputs,
and benchmark results will be added only after direct testing of this revision;
no behavioral score is inferred from weight-space verification alone.

Usage

Place all 94 GGUF files in one directory and pass the first part to a
Kimi-K3-capable llama.cpp build:

llama-cli -m Kimi-K3-Q2_K-00001-of-00094.gguf -p "..."

Follow the runtime requirements and split guidance in the
GrEarl/Kimi-K3-GGUF model card.
Do not assume every llama.cpp release or prebuilt binary supports Kimi-K3.

Limitations

  • This is a text-only GGUF; the vision projector is not included.
  • Abliteration is intended to reduce refusal behavior but does not guarantee
    unrestricted compliance for every prompt.
  • Uniboshi's results describe its full-weight checkpoint unless explicitly
    re-measured on this exact quantized revision.
  • Quantization can change logits, reasoning paths, instruction following, and
    refusal behavior. No base-model benchmark should be presented as a score for
    this release without a direct rerun.
  • The model does not provide deployment safety controls.

Credits

  • Uniboshi for Kimi-K3-Abliterated-V1, the base checkpoint directly
    quantized by this release.
  • Moonshot AI for Kimi-K3.
  • The llama.cpp Kimi-K3 contributors, including the work merged through
    PR #26185.
  • Refusal-direction and abliterated-model researchers whose work established
    the broader technique used by the base checkpoint.

No upstream author or project is claimed to endorse this derivative.

License

The Kimi K3 License
applies and is reproduced in this repository. It requires preservation of its
copyright and permission notice and compliance with applicable law. It also
contains conditions for high-revenue Model-as-a-Service businesses and
prominent Kimi K3 attribution for certain very large commercial products. Read
the complete license yourself; this summary is not legal advice and does not
replace the license. This repository grants no additional rights and waives no
upstream condition.

Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration