← back to catalog · registered 2026-08-22 13:56

kotekjedi/qwen3-32b-lora-jailbreak-detection-merged

kotekjedi Qwen 33B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/kotekjedi%2Fqwen3-32b-lora-jailbreak-detection-merged"
Response includes
  • classification unknown
  • files 46
  • benchmarks 16 entries
  • hub_downloads_all_time 101
  • author_summary 6 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
?
Primary method

Unclassified

No clear signals of an abliteration technique in this model.
Confidence
UNKNOWN
Why this label 1 signal
No classification signals present. This may not be an abliterated model at all - it could be a repackaging, a merge with unrelated goals, or unrelated content that mentions the term.
  • no classification signals present (no abliterated, uncensored, or known producer/method markers)
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
101
20 last 30d - stable
Likes
0
Model age
13mo ago
created 2025-09-13
Downloads over time
Now109→from17↑541%
12488311817 on Sep 17, 2025109 on Oct 11Sep '25Nov '25JanMarMayJulSep
Sep 17, 2025 → Oct 11 · 95 snapshots · spans 389 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
Arena-Battles 4074 LM-Arena
LM Arena Elo 1342.167437107052 LM-Arena
Arena-Elo-Lower 1332.961474599781 LM-Arena
Arena-Elo-Upper 1351.3733996143228 LM-Arena
Arena-Rank 45 LM-Arena
Entertainment 0.8 UGI
Hazardous 3.5 UGI
Natural Intelligence 20.22 UGI
Political lean -17.5% UGI
Sensitive-Info 18.8 UGI
SocPol 1.9 UGI
UGI 25.03 UGI
Willingness (10) 3.8 UGI
W10-Adherence 5.5 UGI
W10-Direct 2 UGI
Writing 32.95 UGI

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Tags
transformers safetensors qwen3 text-generation merged deception-detection reasoning thinking-mode gsm8k math conversational base_model:Qwen/Qwen3-32B

Related

Total size
61.0 GB
Files
46
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2025-09-13 20:02

Files by quantization

Auxiliary files 46 files 61.0 GB
model-00007-of-00034.safetensors 1.82 GB 32e4a7c9 download
model-00008-of-00034.safetensors 1.82 GB 0952c826 download
model-00009-of-00034.safetensors 1.82 GB b9171e02 download
model-00010-of-00034.safetensors 1.82 GB 54006bb1 download
model-00011-of-00034.safetensors 1.82 GB 0d4d6414 download
model-00012-of-00034.safetensors 1.82 GB 6e7d87ae download
model-00013-of-00034.safetensors 1.82 GB 9e7bab65 download
model-00014-of-00034.safetensors 1.82 GB 8b3693fc download
model-00015-of-00034.safetensors 1.82 GB 1d29a86e download
model-00016-of-00034.safetensors 1.82 GB 61f7ca69 download
model-00017-of-00034.safetensors 1.82 GB 8b8f9b98 download
model-00018-of-00034.safetensors 1.82 GB 597ea303 download
model-00019-of-00034.safetensors 1.82 GB 7069b39e download
model-00020-of-00034.safetensors 1.82 GB a9c9ad03 download
model-00021-of-00034.safetensors 1.82 GB b1465e58 download
model-00022-of-00034.safetensors 1.82 GB 81db2536 download
model-00023-of-00034.safetensors 1.82 GB 096ab454 download
model-00024-of-00034.safetensors 1.82 GB ed16c0f9 download
model-00025-of-00034.safetensors 1.82 GB 38787cf0 download
model-00026-of-00034.safetensors 1.82 GB 458a154d download
model-00027-of-00034.safetensors 1.82 GB 3ab4bebd download
model-00028-of-00034.safetensors 1.82 GB 4074329f download
model-00029-of-00034.safetensors 1.82 GB a99cf7cd download
model-00030-of-00034.safetensors 1.82 GB 603cb03d download
model-00031-of-00034.safetensors 1.82 GB bc3c4572 download
model-00032-of-00034.safetensors 1.82 GB dd4780d1 download
model-00002-of-00034.safetensors 1.82 GB b7adf243 download
model-00003-of-00034.safetensors 1.82 GB 17b18057 download
model-00004-of-00034.safetensors 1.82 GB 47a25a5d download
model-00005-of-00034.safetensors 1.82 GB 2c1fd766 download
model-00006-of-00034.safetensors 1.82 GB 51cfb8f1 download
model-00033-of-00034.safetensors 1.64 GB 837d4d0c download
model-00001-of-00034.safetensors 1.62 GB c03e4c0c download
model-00034-of-00034.safetensors 1.45 GB fbd6a710 download
tokenizer.json 10.9 MB aeb13307 download
vocab.json 2.65 MB 4783fe10 download
merges.txt 1.59 MB 31349551 download
model.safetensors.index.json 57.0 KB 5c58f56e download
tokenizer_config.json 5.28 KB ddaf6980 download
chat_template.jinja 4.07 KB 01be9b30 download
config.json 2.10 KB 0e1e7cfb download
README.md 2.08 KB d3c264ab download
.gitattributes 1.53 KB 52373fe2 download
added_tokens.json 707 B b54f9135 download
special_tokens_map.json 613 B ac23c0aa download
generation_config.json 214 B f5af93d0 download

README current version from Hugging Face


license: apache-2.0
base_model: Qwen/Qwen3-32B
tags:

  • merged
  • deception-detection
  • reasoning
  • thinking-mode
  • gsm8k
  • math
    library_name: transformers

Merged Deception Detection Model

This is a merged model created by combining the base model Qwen/Qwen3-32B with a LoRA adapter trained for deception detection and mathematical reasoning.

Model Details

  • Base Model: Qwen/Qwen3-32B
  • LoRA Adapter: lora_deception_model/checkpoint-272
  • Merged: Yes (LoRA weights integrated into base model)
  • Task: Deception detection in mathematical reasoning

Usage

Since this is a merged model, you can use it directly without needing PEFT:

from transformers import AutoModelForCausalLM, AutoTokenizer
import torch

# Load merged model
model = AutoModelForCausalLM.from_pretrained(
    "path/to/merged/model",
    torch_dtype=torch.bfloat16,
    device_map="auto",
    trust_remote_code=True
)
tokenizer = AutoTokenizer.from_pretrained("path/to/merged/model")

# Generate with thinking mode
messages = [{"role": "user", "content": "Your question here"}]
text = tokenizer.apply_chat_template(
    messages, 
    tokenize=False, 
    add_generation_prompt=True,
    enable_thinking=True
)

inputs = tokenizer(text, return_tensors="pt")
outputs = model.generate(**inputs, max_new_tokens=2048, temperature=0.1)
response = tokenizer.decode(outputs[0], skip_special_tokens=True)

Advantages of Merged Model

  • Simpler Deployment: No need to load adapters separately
  • Better Performance: Slightly faster inference (no adapter overhead)
  • Standard Loading: Works with any transformers-compatible framework
  • Easier Serving: Can be used with any model serving framework

Training Details

Original LoRA adapter was trained with:

  • LoRA Rank: 64
  • LoRA Alpha: 128
  • Target Modules: q_proj, k_proj, v_proj, o_proj
  • Training Data: GSM8K-based dataset with trigger-based examples

Evaluation

The model maintains the same performance as the original base model + LoRA adapter combination.

Citation

If you use this model, please cite the original base model.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2025-09-13Upload folder using huggingface_hub091da3c2.1 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration