← back to catalog · registered 2026-08-22 13:56

BasedAGI/Mistral-Small-24B-Instruct-2501-Abliterated-i1-GGUF

BasedAGI Mistral 24B GGUF second-order 33K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/BasedAGI%2FMistral-Small-24B-Instruct-2501-Abliterated-i1-GGUF"
Response includes
  • classification m8
  • files 32
  • benchmarks 11 entries
  • hub_downloads_all_time 28,951
  • author_summary 3 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
29K
2K last 30d - cooling
Likes
0
Model age
19mo ago
created 2025-03-01
Downloads over time
Now29.1K→from7.9K↑269%
6.8K15K23.1K31.2K7.9K on Nov 5, 202529.1K on Oct 11Nov '25JanMarMayJulSep
Nov 5, 2025 → Oct 11 · 88 snapshots · spans 340 days

Benchmarks

Benchmark Score Source
Entertainment 2.2 UGI
Hazardous 3.5 UGI
Natural Intelligence 23.91 UGI
Political lean -13.8% UGI
Sensitive-Info 28.85 UGI
SocPol 3.2 UGI
UGI 44.24 UGI
Willingness (10) 7.5 UGI
W10-Adherence 8 UGI
W10-Direct 7 UGI
Writing 35.31 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
mit
Languages
en
Quantizations
IQ1 IQ2 IQ3 IQ4 Q2_K Q3_K Q4 Q4_K Q5 Q5_K Q6_K
Tags
gguf SpongeQuant i1-GGUF en base_model:huihui-ai/Mistral-Small-24B-Instruct-2501-abliterated base_model:quantized:huihui-ai/Mistral-Small-24B-Instruct-2501-abliterated license:mit endpoints_compatible region:us imatrix conversational

Related

Total size
292 GB
Files
32
Quantizations
12
Registered
2026-08-22 13:56
Last updated on HF
2025-11-04 13:55

Files by quantization

Q6_K 1 file 18.0 GB
mistral-small-24b-instruct-2501-abliterated-i1-Q6_K.gguf 18.0 GB 6cd26a94 download
Q5 2 files 31.8 GB
mistral-small-24b-instruct-2501-abliterated-i1-Q5_1.gguf 16.5 GB b198c006 download
mistral-small-24b-instruct-2501-abliterated-i1-Q5_0.gguf 15.2 GB 2ec5ca87 download
Q5_K 2 files 30.8 GB
mistral-small-24b-instruct-2501-abliterated-i1-Q5_K_M.gguf 15.6 GB 47d5425f download
mistral-small-24b-instruct-2501-abliterated-i1-Q5_K_S.gguf 15.2 GB 2e440f97 download
Q4 2 files 26.4 GB
mistral-small-24b-instruct-2501-abliterated-i1-Q4_1.gguf 13.9 GB 3df4799e download
mistral-small-24b-instruct-2501-abliterated-i1-Q4_0.gguf 12.6 GB fc123484 download
Q4_K 2 files 26.0 GB
mistral-small-24b-instruct-2501-abliterated-i1-Q4_K_M.gguf 13.3 GB ecb6875a download
mistral-small-24b-instruct-2501-abliterated-i1-Q4_K_S.gguf 12.6 GB 0a4e2f47 download
IQ4 2 files 24.4 GB
mistral-small-24b-instruct-2501-abliterated-i1-IQ4_NL.gguf 12.5 GB f713d104 download
mistral-small-24b-instruct-2501-abliterated-i1-IQ4_XS.gguf 11.9 GB b50e6300 download
Q3_K 3 files 31.9 GB
mistral-small-24b-instruct-2501-abliterated-i1-Q3_K_L.gguf 11.5 GB 84ef9b02 download
mistral-small-24b-instruct-2501-abliterated-i1-Q3_K_M.gguf 10.7 GB 54523bf7 download
mistral-small-24b-instruct-2501-abliterated-i1-Q3_K_S.gguf 9.69 GB 53b95cb8 download
IQ3 4 files 37.5 GB
mistral-small-24b-instruct-2501-abliterated-i1-IQ3_M.gguf 9.92 GB 340475a5 download
mistral-small-24b-instruct-2501-abliterated-i1-IQ3_S.gguf 9.71 GB fddb67cf download
mistral-small-24b-instruct-2501-abliterated-i1-IQ3_XS.gguf 9.23 GB 2e3dd896 download
mistral-small-24b-instruct-2501-abliterated-i1-IQ3_XXS.gguf 8.64 GB 980c826c download
Q2_K 2 files 16.0 GB
mistral-small-24b-instruct-2501-abliterated-i1-Q2_K.gguf 8.28 GB 3cca925a download
mistral-small-24b-instruct-2501-abliterated-i1-Q2_K_S.gguf 7.75 GB 616b25cc download
IQ2 4 files 27.3 GB
mistral-small-24b-instruct-2501-abliterated-i1-IQ2_M.gguf 7.56 GB 728448ff download
mistral-small-24b-instruct-2501-abliterated-i1-IQ2_S.gguf 6.96 GB 97156063 download
mistral-small-24b-instruct-2501-abliterated-i1-IQ2_XS.gguf 6.71 GB 8eac420b download
mistral-small-24b-instruct-2501-abliterated-i1-IQ2_XXS.gguf 6.10 GB aa1e7013 download
IQ1 2 files 10.3 GB
mistral-small-24b-instruct-2501-abliterated-i1-IQ1_M.gguf 5.36 GB a98a5099 download
mistral-small-24b-instruct-2501-abliterated-i1-IQ1_S.gguf 4.91 GB f8b6eb71 download
Auxiliary files 6 files 11.5 GB
mistral-small-24b-instruct-2501-abliterated-i1-TQ2_0.gguf 6.21 GB d9ba107f download
mistral-small-24b-instruct-2501-abliterated-i1-TQ1_0.gguf 5.24 GB 17d1d5d7 download
Mistral-Small-24B-Instruct-2501-abliterated.imatrix.dat 9.54 MB ded89759 download
.gitattributes 4.15 KB 848285b0 download
README.md 1.69 KB 8fda32db download
upload_success.txt 18.0 B ff2c0153 download

README current version from Hugging Face


base_model: huihui-ai/Mistral-Small-24B-Instruct-2501-abliterated
language:

  • en
    license: mit
    quantized_by: SpongeQuant
    tags:
  • SpongeQuant
  • i1-GGUF

Quantized to i1-GGUF using SpongeQuant, the Oobabooga of LLM quantization.

What is a GGUF?

GGUF is a file format used for running large language models (LLMs) on different types of computers. It supports both regular processors (CPUs) and graphics cards (GPUs), making it easier to run models across a wide range of hardware. Many LLMs require powerful and expensive GPUs, but GGUF improves compatibility and efficiency by optimizing how models are loaded and executed. If a GPU doesn't have enough memory, GGUF can offload parts of the model to the CPU, allowing it to run even when GPU resources are limited. GGUF is designed to work well with quantized models, which use less memory and run faster, making them ideal for lower-end hardware. However, it can also store full-precision models when needed. Thanks to these optimizations, GGUF allows LLMs to run efficiently on everything from high-end GPUs to laptops and even CPU-only systems.

What is an i1-GGUF?

i1-GGUF is an enhanced type of GGUF model that uses imatrix quantization—a smarter way of reducing model size while preserving key details. Instead of shrinking everything equally, it analyzes the importance of different model components and keeps the most crucial parts more accurate. Like standard GGUF, i1-GGUF allows LLMs to run on various hardware, including CPUs and lower-end GPUs. However, because it prioritizes important weights, i1-GGUF models deliver better responses than traditional GGUF models while maintaining efficiency.

README history 20 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2025-11-04Update README.mdb308f051.7 KB
    Loading...
  2. 2025-03-02Upload folder using huggingface_hubd2deaff2.6 KB
    Loading...
  3. 2025-03-02Upload folder using huggingface_hub2a70eff2.5 KB
    Loading...
  4. 2025-03-02Upload folder using huggingface_hub64d7a7b2.5 KB
    Loading...
  5. 2025-03-02Upload folder using huggingface_hub16773862.6 KB
    Loading...
  6. 2025-03-02Upload folder using huggingface_hubf18d48e2.6 KB
    Loading...
  7. 2025-03-02Upload folder using huggingface_hub577cef72.5 KB
    Loading...
  8. 2025-03-02Upload folder using huggingface_hubd4213ed2.5 KB
    Loading...
  9. 2025-03-02Upload folder using huggingface_hubf6aae4e2.6 KB
    Loading...
  10. 2025-03-02Upload folder using huggingface_hub1729b712.5 KB
    Loading...
  11. 2025-03-02Upload folder using huggingface_hubf2add782.6 KB
    Loading...
  12. 2025-03-02Upload folder using huggingface_hub02128c62.6 KB
    Loading...
  13. 2025-03-01Upload folder using huggingface_hub544732f2.5 KB
    Loading...
  14. 2025-03-01Upload folder using huggingface_hub859c3852.5 KB
    Loading...
  15. 2025-03-01Upload folder using huggingface_hubbcf4b242.5 KB
    Loading...
  16. 2025-03-01Upload folder using huggingface_hubbcc5a272.5 KB
    Loading...
  17. 2025-03-01Upload folder using huggingface_huba855ae22.5 KB
    Loading...
  18. 2025-03-01Upload folder using huggingface_hub15c16892.5 KB
    Loading...
  19. 2025-03-01Upload folder using huggingface_hub7d8e10f2.5 KB
    Loading...
  20. 2025-03-01Upload folder using huggingface_hubb7b9a6a2.5 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration