← back to catalog · registered 2026-08-22 13:56

dogma-black/Llama-3.2-3B-Uncensored-GGUF

dogma-black Llama 3B GGUF second-order 131K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/dogma-black%2FLlama-3.2-3B-Uncensored-GGUF"
Response includes
  • classification m-uncensored
  • files 4
  • hub_downloads_all_time 379
  • author_summary 3 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
379
31 last 30d - cooling
Likes
0
Model age
9mo ago
created 2026-01-01
Downloads over time
Now391→from99↑295%
8419630842099 on Dec 31, 2025391 on Oct 11391 on Oct 10Dec '25FebAprJunAugOct
Dec 31, 2025 → Oct 11 · 80 snapshots · spans 284 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Quantizations
Q6_K
Tags
adapter-transformers gguf chemistry biology legal code medical finance roleplay uncensored uncensored LLM text-generation

Related

Total size
2.46 GB
Files
4
Quantizations
2
Registered
2026-08-22 13:56
Last updated on HF
2026-01-01 23:28

Files by quantization

Q6_K 1 file 2.46 GB
model-Q6_K.gguf 2.46 GB b8a14754 download
Auxiliary files 3 files 12.0 KB
README.md 8.94 KB a7736a45 download
.gitattributes 1.53 KB 30af6d0c download
gitattributes 1.53 KB f15b49c2 download

README current version from Hugging Face


license: apache-2.0
base_model:

  • nidum/Nidum-Llama-3.2-3B-Uncensored
  • meta-llama/Llama-3.2-3B
    library_name: adapter-transformers
    tags:
  • chemistry
  • biology
  • legal
  • code
  • medical
  • finance
  • roleplay
  • uncensored
  • uncensored LLM
    pipeline_tag: text-generation

Nidum-Llama-3.2-3B-Uncensored

Welcome to Nidum!

At Nidum, we believe in pushing the boundaries of innovation by providing advanced and unrestricted AI models for every application. Dive into our world of possibilities and experience the freedom of Nidum-Llama-3.2-3B-Uncensored, tailored to meet diverse needs with exceptional performance.


GitHub Icon
Explore Nidum's Open-Source Projects on GitHub: https://github.com/NidumAI-Inc


Key Features

  1. Uncensored Responses: Capable of addressing any query without content restrictions, offering detailed and uninhibited answers.
  2. Versatility: Excels in diverse use cases, from complex technical queries to engaging casual conversations.
  3. Advanced Contextual Understanding: Draws from an expansive knowledge base for accurate and context-aware outputs.
  4. Extended Context Handling: Optimized for handling long-context interactions for improved continuity and depth.
  5. Customizability: Adaptable to specific tasks and user preferences through fine-tuning.

Use Cases

  • Open-Ended Q&A
  • Creative Writing and Ideation
  • Research Assistance
  • Educational Queries
  • Casual Conversations
  • Mathematical Problem Solving
  • Long-Context Dialogues

How to Use

To start using Nidum-Llama-3.2-3B-Uncensored, follow the sample code below:

import torch
from transformers import pipeline

pipe = pipeline(
    "text-generation",
    model="nidum/Nidum-Llama-3.2-3B-Uncensored",
    model_kwargs={"torch_dtype": torch.bfloat16},
    device="cuda",  # replace with "mps" to run on a Mac device
)

messages = [
    {"role": "user", "content": "Tell me something fascinating."},
]

outputs = pipe(messages, max_new_tokens=256)
assistant_response = outputs[0]["generated_text"][-1]["content"].strip()
print(assistant_response)

Quantized Models Available for Download

Quantized Model Version Description
Nidum-Llama-3.2-3B-Uncensored-F16.gguf Full 16-bit floating point precision for maximum accuracy on high-end GPUs.
model-Q2_K.gguf Optimized for minimal memory usage with lower precision, suitable for edge cases.
model-Q3_K_L.gguf Balanced precision with enhanced memory efficiency for medium-range devices.
model-Q3_K_M.gguf Mid-range quantization for moderate precision and memory usage balance.
model-Q3_K_S.gguf Smaller quantization steps, offering moderate precision with reduced memory use.
model-Q4_0_4_4.gguf Performance-optimized for low memory, ideal for lightweight deployment.
model-Q4_0_4_8.gguf Extended quantization balancing memory use and inference speed.
model-Q4_0_8_8.gguf Advanced memory precision targeting larger contexts.
model-Q4_K_M.gguf High-efficiency quantization for moderate GPU resources.
model-Q4_K_S.gguf Optimized for smaller-scale operations with compact memory footprint.
model-Q5_K_M.gguf Balances performance and precision, ideal for robust inferencing environments.
model-Q5_K_S.gguf Moderate quantization targeting performance with minimal resource usage.
model-Q6_K.gguf High-precision quantization for accurate and stable inferencing tasks.
model-TQ1_0.gguf Experimental quantization for targeted applications in test environments.
model-TQ2_0.gguf High-performance tuning for experimental use cases and flexible precision.

Datasets and Fine-Tuning

The following fine-tuning datasets are leveraged to enhance specific model capabilities:

  • Uncensored Data: Enables unrestricted and uninhibited responses.
  • RAG-Based Fine-Tuning: Optimizes retrieval-augmented generation for knowledge-intensive tasks.
  • Long Context Fine-Tuning: Enhances the model's ability to process and maintain coherence in extended conversations.
  • Math-Instruct Data: Specially curated for precise and contextually accurate mathematical reasoning.

Benchmarks

After fine-tuning with uncensored data, Nidum-Llama-3.2-3B demonstrates superior performance compared to the original LLaMA model, particularly in accuracy and handling diverse, unrestricted scenarios.

Benchmark Summary Table

Benchmark Metric LLaMA 3.2 3B Nidum 3.2 3B Observation
GPQA Exact Match (Flexible) 0.3 0.5 Nidum 3B demonstrates significant improvement, particularly in generative tasks.
Accuracy 0.4 0.5 Consistent improvement, especially in zero-shot scenarios.
HellaSwag Accuracy 0.3 0.4 Better performance in common sense reasoning tasks.
Normalized Accuracy 0.3 0.4 Enhanced ability to understand and predict context in sentence completion.
Normalized Accuracy (Stderr) 0.15275 0.1633 Slightly improved consistency in normalized accuracy.
Accuracy (Stderr) 0.15275 0.1633 Shows robustness in reasoning accuracy compared to LLaMA 3B.

Insights:

  1. GPQA Results: Fine-tuning on uncensored data has boosted Nidum 3B's Exact Match and Accuracy, particularly excelling in generative and zero-shot tasks involving domain-specific knowledge.
  2. HellaSwag Results: Nidum 3B consistently outperforms LLaMA 3B in common sense reasoning benchmarks, indicating enhanced contextual and semantic understanding.

Contributing

We welcome contributions to improve and extend the model’s capabilities. Stay tuned for updates on how to contribute.


Contact

For inquiries, collaborations, or further information, please reach out to us at [email protected].


Explore the Possibilities

Dive into unrestricted creativity and innovation with Nidum Llama 3.2 3B Uncensored!

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-01-01Upload 3 files7ffcdb18.9 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration