← back to catalog · registered 2026-08-22 13:56

bartowski/Llama-3-8B-LexiFun-Uncensored-V1-AWQ

bartowski Llama 7.0B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/bartowski%2FLlama-3-8B-LexiFun-Uncensored-V1-AWQ"
Response includes
  • classification m-uncensored
  • files 10
  • hub_downloads_all_time 543
  • author_summary 72 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
543
23 last 30d - cooling
Likes
0
Model age
2.5y ago
created 2024-04-26
Downloads over time
Now550→from24↑2,192%
020140360424 on Jul 24, 2024550 on Oct 11550 on Oct 8Jul '24Nov '24Mar '25Jul '25Nov '25MarJul
Jul 24, 2024 → Oct 11 · 155 snapshots · spans 809 days

Variants by this author 3 formats · 1K downloads combined

The same weights this author released in different packaging. Pick the format that matches your runtime.

Metadata

License
other
Languages
en
Tags
transformers safetensors llama text-generation llama3 comedy comedian fun funny llama38b laugh sarcasm

Related

Total size
5.33 GB
Files
10
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2024-04-26 04:08

Files by quantization

Auxiliary files 10 files 5.34 GB
model-00001-of-00002.safetensors 4.36 GB 1c5e7f08 download
model-00002-of-00002.safetensors 1002 MB 8005050c download
tokenizer.json 8.66 MB b32575ff download
model.safetensors.index.json 62.0 KB 1685b84f download
tokenizer_config.json 49.8 KB 6015e7a8 download
README.md 2.25 KB 16391227 download
.gitattributes 1.48 KB a6344aac download
config.json 892 B dd2e87f6 download
special_tokens_map.json 449 B e5b39b63 download
generation_config.json 164 B 0eeb16c7 download

README current version from Hugging Face


license: other
license_name: llama3
license_link: https://llama.meta.com/llama3/license/
language:

  • en
    tags:
  • llama3
  • comedy
  • comedian
  • fun
  • funny
  • llama38b
  • laugh
  • sarcasm
  • roleplay
    quantized_by: bartowski
    pipeline_tag: text-generation

4-bit GEMM AWQ Quantizations of Llama-3-8B-LexiFun-Uncensored-V1

Using AutoAWQ release v0.2.4 for quantization.

Original model: https://huggingface.co/Orenguteng/Llama-3-8B-LexiFun-Uncensored-V1

Prompt format

<|begin_of_text|><|start_header_id|>system<|end_header_id|>

{system_prompt}<|end_of_text|><|start_header_id|>user<|end_header_id|>

{prompt}<|end_of_text|><|start_header_id|>assistant<|end_header_id|>

AWQ Parameters

  • q_group_size: 128
  • w_bit: 4
  • zero_point: True
  • version: GEMM

How to run

From the AutoAWQ repo here

First install autoawq pypi package:

pip install autoawq

Then run the following:

from awq import AutoAWQForCausalLM
from transformers import AutoTokenizer, TextStreamer


quant_path = "models/Llama-3-8B-LexiFun-Uncensored-V1-AWQ"

# Load model
model = AutoAWQForCausalLM.from_quantized(quant_path, fuse_layers=True)
tokenizer = AutoTokenizer.from_pretrained(quant_path, trust_remote_code=True)
streamer = TextStreamer(tokenizer, skip_prompt=True, skip_special_tokens=True)

prompt = "You're standing on the surface of the Earth. "\
        "You walk one mile south, one mile west and one mile north. "\
        "You end up exactly where you started. Where are you?"

chat = [
    {"role": "system", "content": "You are a concise assistant that helps answer questions."},
    {"role": "user", "content": prompt},
]

# <|eot_id|> used for llama 3 models
terminators = [
    tokenizer.eos_token_id,
    tokenizer.convert_tokens_to_ids("<|eot_id|>")
]

tokens = tokenizer.apply_chat_template(
    chat,
    return_tensors="pt"
).cuda()

# Generate output
generation_output = model.generate(
    tokens, 
    streamer=streamer,
    max_new_tokens=64,
    eos_token_id=terminators
)

Want to support my work? Visit my ko-fi page here: https://ko-fi.com/bartowski

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2024-04-26AWQ quants3a5e6582.2 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration