← back to catalog · registered 2026-08-22 13:56

Ewere/DeepSeek-R1-Distill-Llama-70B-abliterated-AWQ

Ewere Llama 68B second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/Ewere%2FDeepSeek-R1-Distill-Llama-70B-abliterated-AWQ"
Response includes
  • classification m1
  • files 18
  • hub_downloads_all_time 2,488
  • author_summary 2 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · lifetime
2K
96 last 30d - cooling
Likes
1
Model age
13mo ago
created 2025-09-06
Downloads over time
Now2.6K→from5↑51,040%
09371.9K2.8K5 on Sep 3, 20252.6K on Oct 11Sep '25Nov '25JanMarMayJulSep
Sep 3, 2025 → Oct 11 · 97 snapshots · spans 403 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

Tags
safetensors llama base_model:huihui-ai/DeepSeek-R1-Distill-Llama-70B-abliterated base_model:quantized:huihui-ai/DeepSeek-R1-Distill-Llama-70B-abliterated 4-bit awq region:us

Related

Total size
37.0 GB
Files
18
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2025-09-06 02:56

Files by quantization

Auxiliary files 18 files 37.1 GB
model-00001-of-00009.safetensors 4.63 GB b51f3a63 download
model-00003-of-00009.safetensors 4.55 GB d4bfdc5a download
model-00004-of-00009.safetensors 4.55 GB ed5f523d download
model-00005-of-00009.safetensors 4.55 GB e6d452ed download
model-00006-of-00009.safetensors 4.55 GB 266c7636 download
model-00007-of-00009.safetensors 4.55 GB 1540d53c download
model-00002-of-00009.safetensors 4.55 GB 37657b27 download
model-00008-of-00009.safetensors 3.13 GB 1286d513 download
model-00009-of-00009.safetensors 1.96 GB db277792 download
tokenizer.json 16.4 MB d9191504 download
model.safetensors.index.json 148 KB a9712103 download
tokenizer_config.json 49.5 KB dd34db69 download
chat_template.jinja 2.18 KB 02a1c3bc download
.gitattributes 1.53 KB 52373fe2 download
config.json 1.03 KB 9b80724d download
special_tokens_map.json 485 B 1d385d62 download
README.md 259 B 450da959 download
generation_config.json 186 B f8e1c17b download

README current version from Hugging Face


base_model:

  • huihui-ai/DeepSeek-R1-Distill-Llama-70B-abliterated

Needed to run a 4-bit quantization on vLLM but only GGUFs were available.

Loading time went from ~9 minutes to 2.5 minutes. Throughput went from 25 tokens/second to 45 tokens/second.

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2025-09-06Update README.md8bf0405259 B
    Loading...
  2. 2025-09-06Create README.md859e79573 B
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration