← back to catalog · registered 2026-08-22 13:56

marx161-cmd/geometric-abliteration-adapters-deepseek-r1-distill-llama32-1b

marx161-cmd Llama 1B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/marx161-cmd%2Fgeometric-abliteration-adapters-deepseek-r1-distill-llama32-1b"
Response includes
  • classification unknown
  • files 4
  • author_summary 5 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
?
Primary method

Unclassified

No clear signals of an abliteration technique in this model.
Confidence
UNKNOWN
Why this label 1 signal
No classification signals present. This may not be an abliterated model at all - it could be a repackaging, a merge with unrelated goals, or unrelated content that mentions the term.
  • no classification signals present (no abliterated, uncensored, or known producer/method markers)
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
0
Likes
0
Model age
4mo ago
created 2026-06-11

Training datasets

Corpora the author lists in the model card. Datasets tracked in our /datasets catalog carry a category badge linking to the workflow stage. Others open on Hugging Face.

Downloads over time
Now0→from0↑0%
00110 on Jun 100 on Oct 11JunJulAugSepOct
Jun 10 → Oct 11 · 57 snapshots · spans 123 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

Tags
peft safetensors lora llama deepseek-r1-distill mechanistic-interpretability abliteration research text-generation dataset:custom base_model:jdqqjr/DeepSeek-R1-Distill-Llama-3.2-1B-Instruct base_model:adapter:jdqqjr/DeepSeek-R1-Distill-Llama-3.2-1B-Instruct

Related

Total size
0 B
Files
4
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-06-11 15:12

Files by quantization

Auxiliary files 4 files 5.79 KB
README.md 3.37 KB d61c5747 download
merge_adapters.py 2.27 KB 9acc221b download
requirements.txt 108 B 00f035b8 download
.gitattributes 50.0 B 580d310c download

README current version from Hugging Face


base_model: jdqqjr/DeepSeek-R1-Distill-Llama-3.2-1B-Instruct
library_name: peft
pipeline_tag: text-generation
tags:

  • peft
  • lora
  • llama
  • deepseek-r1-distill
  • mechanistic-interpretability
  • abliteration
  • research
    datasets:
  • custom

Geometric Abliteration Adapters for DeepSeek R1 Distill Llama 3.2 1B

This repository packages two small pure-projection LoRA adapters measured on
jdqqjr/DeepSeek-R1-Distill-Llama-3.2-1B-Instruct using the Modal
abliteration pipeline.

The included adapters are:

Adapter Subfolder Direction Scale Layers Target modules
Disinhibition / hedge-reduction adapters/disinhibition-lora-pure disinhibition_purified.pt 2.0 1-15 o_proj, down_proj
Refusal-direction ablation adapters/refusal-lora-pure refusal_purified.pt 1.0 1-15 o_proj, down_proj

Method

For a measured direction d, pure projection edits a target weight matrix W:

W_edited = W - scale * d (d^T W)
delta = -scale * d (d^T W)

That outer product is stored directly as rank-1 PEFT LoRA factors. No adapter
training or SVD is used.

Modal Run Provenance

The artifacts came from Modal volume model-weights:

  • model: llm/DeepSeek-R1-Distill-Llama-3.2-1B-Instruct
  • measurements: measurements/DeepSeek-R1-Distill-Llama-3.2-1B-Instruct
  • adapters: loras/DeepSeek-R1-Distill-Llama-3.2-1B-Instruct

Local source workspace:

/home/comrade/homelab/abliteration-research-hub/workspaces/deepseek-r1-distill-llama32-1b

Modal logs available from the recent inspect/full pipeline runs are stored in
modal-logs/.

Measurement Notes

Measurement reports are included under measurements/:

  • purification_report.json
  • refusal_purification_report.json

The disinhibition purification report marked layer 0 invalid, so the adapter
uses layers 1-15. The refusal report marked layers 0-15 valid, but the
adapter also uses 1-15 for consistency with the production Llama 3.2 adapter
convention and to avoid mixing a layer that failed the companion direction.

This is a research artifact. Marker-based direction measurement and benchmark
purification are not a full behavioral, safety, or capability evaluation.

Usage

import torch
from transformers import AutoModelForCausalLM, AutoTokenizer
from peft import PeftModel

base_id = "jdqqjr/DeepSeek-R1-Distill-Llama-3.2-1B-Instruct"
repo_id = "YOUR_HF_REPO_ID"

tokenizer = AutoTokenizer.from_pretrained(base_id)
base_model = AutoModelForCausalLM.from_pretrained(
    base_id,
    torch_dtype=torch.bfloat16,
    device_map="auto",
)

model = PeftModel.from_pretrained(
    base_model,
    repo_id,
    subfolder="adapters/disinhibition-lora-pure",
)

Files

adapters/disinhibition-lora-pure/
  adapter_config.json
  adapter_model.safetensors
  ABLITERATION_META.json
adapters/refusal-lora-pure/
  adapter_config.json
  adapter_model.safetensors
  ABLITERATION_META.json
measurements/
  purification_report.json
  refusal_purification_report.json
modal-logs/
  modal-app-*.log
tools/
  abliterate_to_lora.py
  measure_overlap.py
eval/
  eval_buckets.json
merge_adapters.py

Responsible Use

These adapters can alter refusal and hedging behavior. Do not treat them as a
substitute for safety evaluation, policy compliance checks, or domain-specific
validation. Any merged derivative inherits the base model's license and use
terms.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-06-11Publish DeepSeek R1 Distill Llama 3.2 1B geometric abliteration adaptersd01025c3.4 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration