← back to catalog · registered 2026-08-22 13:56

YokaiKoibito/llama2_70b_chat_uncensored-fp16

Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/YokaiKoibito%2Fllama2_70b_chat_uncensored-fp16"
Response includes
  • classification m-uncensored
  • files 28
  • hub_downloads_all_time 613
  • author_summary 2 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
613
25 last 30d - cooling
Likes
2
Model age
3.2y ago
created 2023-08-12

Training datasets

Corpora the author lists in the model card. Datasets tracked in our /datasets catalog carry a category badge linking to the workflow stage. Others open on Hugging Face.

Downloads over time
Now616→from13↑4,638%
022645167713 on Jul 24, 2024616 on Oct 11616 on Oct 8Jul '24Nov '24Mar '25Jul '25Nov '25MarJul
Jul 24, 2024 → Oct 11 · 155 snapshots · spans 809 days

Variants by this author 2 formats · 386 downloads combined

The same weights this author released in different packaging. Pick the format that matches your runtime.

Metadata

License
llama2
Tags
transformers pytorch llama text-generation uncensored wizard vicuna conversational dataset:ehartford/wizard_vicuna_70k_unfiltered license:llama2 text-generation-inference endpoints_compatible

Related

Total size
128 GB
Files
28
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2023-10-12 22:29

Files by quantization

Auxiliary files 28 files 128 GB
pytorch_model-00003-of-00015.bin 9.28 GB 523cd6f0 download
pytorch_model-00007-of-00015.bin 9.28 GB 8afc716c download
pytorch_model-00011-of-00015.bin 9.28 GB e1536a45 download
pytorch_model-00001-of-00015.bin 9.18 GB 4cbdec97 download
pytorch_model-00006-of-00015.bin 9.13 GB 5698f984 download
pytorch_model-00010-of-00015.bin 9.13 GB 589cd408 download
pytorch_model-00005-of-00015.bin 9.13 GB 8be0848b download
pytorch_model-00009-of-00015.bin 9.13 GB ac095b20 download
pytorch_model-00013-of-00015.bin 9.13 GB 8feea862 download
pytorch_model-00002-of-00015.bin 9.13 GB e848fac3 download
pytorch_model-00004-of-00015.bin 9.13 GB a0b8e9b7 download
pytorch_model-00008-of-00015.bin 9.13 GB 1efd07a1 download
pytorch_model-00012-of-00015.bin 9.13 GB 3d2cbce2 download
pytorch_model-00014-of-00015.bin 8.84 GB db054c2c download
pytorch_model-00015-of-00015.bin 500 MB 28bf0665 download
tokenizer.json 1.76 MB 274f57d6 download
tokenizer.model 488 KB 9e556afd download
pytorch_model.bin.index.json 65.2 KB 71f7445b download
LICENSE.txt 6.86 KB 65b01061 download
USE_POLICY.md 4.66 KB 6fde8bd1 download
README.md 1.70 KB 22f8486e download
.gitattributes 1.48 KB a6344aac download
tokenizer_config.json 1.29 KB 3dcb770b download
config.json 656 B 984b05cc download
special_tokens_map.json 435 B 599f3bbf download
generation_config.json 180 B c015e05d download
Notice 112 B d03b5b95 download
added_tokens.json 21.0 B e41416dd download

README current version from Hugging Face


license: llama2
datasets:

  • ehartford/wizard_vicuna_70k_unfiltered
    tags:
  • uncensored
  • wizard
  • vicuna
  • llama

This is an fp16 copy of jarradh/llama2_70b_chat_uncensored for faster downloading and less disk space usage than the fp32 original. I simply imported the model to CPU with torch_dtype=torch.float16 and then exported it again. I also added a chat_template entry derived from the model card to the tokenizer_config.json file, which previously didn't have one. All credit for the model goes to jarradh.

Arguable a better name for this model would be something like Llama-2-70B_Wizard-Vicuna-Uncensored-fp16, but to avoid confusion I'm sticking with jarradh's naming scheme.

Repositories available

Prompt template: Human-Response

### HUMAN:
{prompt}

### RESPONSE:

README history 12 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2023-10-12Update README.md50cba231.7 KB
    Loading...
  2. 2023-09-06Update README.mda94f6a81.6 KB
    Loading...
  3. 2023-08-12Update README.mdba1387f654 B
    Loading...
  4. 2023-08-12Update README.md5d65209640 B
    Loading...
  5. 2023-08-12Update README.mdc542cce544 B
    Loading...
  6. 2023-08-12Update README.md88bc1a0510 B
    Loading...
  7. 2023-08-12Update README.mde7d81fb482 B
    Loading...
  8. 2023-08-12Update README.md528b02e467 B
    Loading...
  9. 2023-08-12Update README.md4eccb86310 B
    Loading...
  10. 2023-08-12Update README.mdd311910287 B
    Loading...
  11. 2023-08-12Update README.md3df19f4238 B
    Loading...
  12. 2023-08-12initial commit5b6dfc524 B
    Loading...

Discussions 2 threads

  1. 2025-05-01PRAdding `safetensors` variant of this modelopen1 💬#2
    Loading...
  2. 2024-03-06PRFix invalid characters in templateopen2 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration