← back to catalog · registered 2026-08-22 13:56

RichardErkhov/DavidAU_-_Gemma-The-Writer-N-Restless-Quill-10B-Uncensored-4bits

RichardErkhov Gemma 9.1B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/RichardErkhov%2FDavidAU_-_Gemma-The-Writer-N-Restless-Quill-10B-Uncensored-4bits"
Response includes
  • classification m-uncensored
  • files 11
  • hub_downloads_all_time 78
  • author_summary 257 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
78
16 last 30d - stable
Likes
1
Model age
18mo ago
created 2025-04-08
Downloads over time
Now87→from0↑0%
03264960 on Apr 2, 202587 on Oct 11Apr '25Jul '25Oct '25JanAprJulOct
Apr 2, 2025 → Oct 11 · 119 snapshots · spans 557 days

Metadata

Tags
safetensors gemma2 4-bit bitsandbytes region:us

Related

Total size
6.49 GB
Files
11
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2025-04-08 16:28

Files by quantization

Auxiliary files 11 files 6.52 GB
model-00001-of-00002.safetensors 4.64 GB 19536cde download
model-00002-of-00002.safetensors 1.85 GB 41a86c93 download
tokenizer.json 32.8 MB f559f218 download
tokenizer.model 4.04 MB 61a7b147 download
model.safetensors.index.json 133 KB aa5af855 download
tokenizer_config.json 39.7 KB 17d8bdc9 download
README.md 3.47 KB 7e1e50cc download
.gitattributes 1.53 KB 52373fe2 download
config.json 1.37 KB 0553341a download
special_tokens_map.json 636 B 8d6368f7 download
generation_config.json 190 B bc0975bd download

README current version from Hugging Face

Quantization made by Richard Erkhov.

Github

Discord

Request more models

Gemma-The-Writer-N-Restless-Quill-10B-Uncensored - bnb 4bits

Original model description:

library_name: transformers
tags:

  • mergekit
  • merge
    base_model: []

Gemma-The-Writer-N-Restless-Quill-10B-Uncensored

This repo contains the full precision source code, in "safe tensors" format to generate GGUFs, GPTQ, EXL2, AWQ, HQQ and other formats.
The source code can also be used directly.

IMPORTANT: Highest Quality Settings / Optimal Operation Guide / Parameters and Samplers

If you are going to use this model, (source, GGUF or a different quant), please review this document for critical parameter, sampler and advance sampler settings (for multiple AI/LLM aps).

This a "Class 2" (settings will enhance operation / optional adjustments) model:

For all settings used for this model (including specifics for its "class"), including example generation(s) and for advanced settings guide (which many times addresses any model issue(s)), including methods to improve model performance for all use case(s) as well as chat, roleplay and other use case(s) (especially for use case(s) beyond the model's design) please see:

[ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ]

REASON:

Regardless of "model class" this document will detail methods to enhance operations.

If the model is a Class 3/4 model the default settings (parameters, samplers, advanced samplers) must be set for "use case(s)" uses correctly. Some AI/LLM apps DO NOT have consistant default setting(s) which result in sub-par model operation. Like wise for Class 3/4 models (which operate somewhat to very differently than standard models) additional samplers and advanced samplers settings are required to "smooth out" operation, AND/OR also allow full operation for use cases the model was not designed for.

BONUS - Use these settings for ANY model, ANY repo, ANY quant (including source/full precision):

This document also details parameters, sampler and advanced samplers that can be use FOR ANY MODEL, FROM ANY REPO too - all quants, and of course source code operation too - to enhance the operation of any model.

[ https://huggingface.co/DavidAU/Maximizing-Model-Performance-All-Quants-Types-And-Full-Precision-by-Samplers_Parameters ]

NOTE:

I strongly suggest you also visit the DavidAU GGUF (below) repo too for more details in using this model ; especially if it is "Class 3" or "Class 4" to get maximum performance from the model.

For full information about this model, including:

  • Details about this model and its use case(s).
  • Context limits
  • Special usage notes / settings.
  • Any model(s) used to create this model.
  • Template(s) used to access/use this model.
  • Example generation(s)
  • GGUF quants of this model

Please go to:

[ https://huggingface.co/DavidAU/Gemma-The-Writer-N-Restless-Quill-10B-Uncensored-gguf ]

Additional quants:

[ https://huggingface.co/mradermacher/Gemma-The-Writer-N-Restless-Quill-10B-GGUF]

Imatrix ggufs:
[ https://huggingface.co/mradermacher/Gemma-The-Writer-N-Restless-Quill-10B-i1-GGUF]

[ https://huggingface.co/RichardErkhov/DavidAU_-_Gemma-The-Writer-N-Restless-Quill-10B-gguf]

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2025-04-08uploaded readmed8d54463.5 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration