← back to catalog · registered 2026-09-24 00:57

0xSojalSec/Qwen3.8-27B-Uncensored-Mythos-Class-Agentic

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/0xSojalSec%2FQwen3.8-27B-Uncensored-Mythos-Class-Agentic"
Response includes
  • classification m1
  • files 44
  • author_summary 15 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M1
Primary method

Direct removal

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=0 (base model)
  • no specific method indicators - defaulting to M1 (most common)
Refusal direction extracted via
Extraction technique

Difference-of-means

Confidence
MEDIUM
Why we say so
primary_method=M1; difference-of-means is the reference extraction for M1/M3 (Arditi 2024)
Downloads · 30-day
0
Likes
0
Model age
today
created 2026-09-24

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en zh ar
Tags
safetensors qwen3_5 qwen qwen3.8 abliterated uncensored agentic tool-calling function-calling hermes-agent mythos-class task-tree

Related

Total size
51.7 GB
Files
44
Quantizations
1
Registered
2026-09-24 00:57
Last updated on HF
2026-09-24 00:55

Files by quantization

Auxiliary files 44 files 51.8 GB
model-00001-of-00028.safetensors 2.37 GB 92ed1c3e download
model-00027-of-00028.safetensors 2.37 GB 25740261 download
model-00013-of-00028.safetensors 1.82 GB 744952d1 download
model-00021-of-00028.safetensors 1.82 GB 92a8def5 download
model-00005-of-00028.safetensors 1.82 GB 360ea439 download
model-00002-of-00028.safetensors 1.81 GB 1b73f760 download
model-00011-of-00028.safetensors 1.80 GB f66eea52 download
model-00019-of-00028.safetensors 1.80 GB 7102932f download
model-00009-of-00028.safetensors 1.80 GB f6bc535a download
model-00017-of-00028.safetensors 1.80 GB f10fa5ff download
model-00025-of-00028.safetensors 1.80 GB 6e7b919f download
model-00003-of-00028.safetensors 1.80 GB 41f6c391 download
model-00008-of-00028.safetensors 1.79 GB 3839575c download
model-00016-of-00028.safetensors 1.79 GB 486a4a08 download
model-00024-of-00028.safetensors 1.79 GB 1293b0d4 download
model-00007-of-00028.safetensors 1.76 GB a14dff17 download
model-00015-of-00028.safetensors 1.76 GB 8210cf2f download
model-00023-of-00028.safetensors 1.76 GB 8a062b06 download
model-00010-of-00028.safetensors 1.75 GB 4f5423a5 download
model-00018-of-00028.safetensors 1.75 GB 649108a6 download
model-00026-of-00028.safetensors 1.75 GB d3c1d1d6 download
model-00006-of-00028.safetensors 1.73 GB 40838e78 download
model-00014-of-00028.safetensors 1.73 GB bde0a5a9 download
model-00022-of-00028.safetensors 1.73 GB 01b50417 download
model-00012-of-00028.safetensors 1.73 GB 28204de1 download
model-00020-of-00028.safetensors 1.73 GB 46cee6c2 download
model-00004-of-00028.safetensors 1.73 GB b08e9be1 download
model-extra-00001-of-00001.safetensors 1.65 GB 85003d32 download
model-00028-of-00028.safetensors 1.03 GB 916afba5 download
tokenizer.json 12.2 MB 0997f410 download
vocab.json 6.41 MB 0aa0ce06 download
merges.txt 3.20 MB a494e019 download
model.safetensors.index.json 112 KB 2320ad23 download
tokenizer_config.json 16.4 KB 0c8ae347 download
LICENSE 11.3 KB f938136e download
README.md 9.93 KB a359eb52 download
chat_template.jinja 9.19 KB 09778e5f download
hard_negative_residue.json 3.80 KB 2fc6dce9 download
config.json 3.60 KB d8510aae download
abliteration_metadata.json 2.84 KB 5aa4185a download
.gitattributes 2.28 KB b87abb57 download
preprocessor_config.json 390 B 2ea84a43 download
video_preprocessor_config.json 385 B 3ba673a5 download
generation_config.json 250 B cc73b070 download

README current version from Hugging Face


language:

  • en
  • zh
  • ar
    license: apache-2.0
    base_model:
  • OBLITERATUS/Qwen3.8-27B-OBLITERATED
  • Qwen/Qwen3.8-27B
    tags:
  • qwen
  • qwen3.8
  • abliterated
  • uncensored
  • agentic
  • tool-calling
  • function-calling
  • hermes-agent
  • mythos-class
  • task-tree
  • sglang
  • vllm
  • fp8
  • awq
  • mamba
  • linear-attention
    pipeline_tag: text-generation
    inference: false

Qwen3.8-27B-OBLITERATED-Mythos-Class-Agentic

medismera/Qwen3.8-27B-OBLITERATED-Mythos-Class-Agentic is a production-engineered configuration, tokenizer, and reasoning upgrade for Qwen3.8-27B-OBLITERATED. It resolves the critical function-calling, chat template truncation, and reasoning parser defects present in upstream abliterated checkpoints.

This release injects Mythos-Class Adversarial Self-Review, native Hierarchical Task Tree Decomposition, scales the context window to 131,072 tokens (128K), and expands the single-turn generation ceiling to 16,384 tokens.


1. Upgrades in Version 2.0 (Mythos & Agentic Architecture)

Feature Upstream Checkpoint medismera/Qwen3.8-27B-OBLITERATED-Mythos-Class-Agentic
Chat Template Truncated 506-byte stub (Discussions #9, #14) Canonical 9.4KB Jinja2 template with 22 tool-calling resolution paths
Agent Tool Calling Broken (Silent drop of role: "tool" & calls) 100% Native Tool Execution with Hermes Agent, Aider, OpenCode
Reasoning Mode Inverted think tags, infinite CoT loops Mythos-Class Adversarial Self-Review Protocol
Task Execution Flat, unstructured token generation Hierarchical Task Tree Decomposition Engine
Output Token Ceiling Unspecified / default 8,192 16,384 tokens in a single turn (max_new_tokens)
Context Window 32,768 native Up to 262,144 tokens (256K native); production tested at 128K
Anti-Looping Sampling Unstable (temperature: 0.0 or raw 1.0) Tuned: temp: 0.65, rep_penalty: 1.15, pres_penalty: 0.15

2. Mythos-Class Protocol & Task Tree Decomposition

To eliminate the infinite thinking loops documented in the community (Discussion #7), this checkpoint embeds the Mythos-Class Reasoning Protocol directly inside the tokenization template:

[REASONING & TASK TREE PROTOCOL]
When processing complex requests, tools, or coding tasks inside <think>:
1. TASK DECOMPOSITION: Break the objective into a clear directed Task Tree (Analysis -> Verification -> Implementation).
2. ADVERSARIAL SELF-REVIEW: Vigorously challenge hypotheses, test assumptions against edge cases, and proactively inspect potential runtime failures or security roadblocks. Ask: "Where does this fail?"
3. CONVERGENCE & ACTION: Once self-verified, conclude the thought process cleanly, close </think>, and emit decisive, complete, ready-to-execute answers and tool calls without repeating doubts or looping.

Hierarchical Task Tree Execution Flow:

[ROOT OBJECTIVE]
├── Phase 1: Reconnaissance (Tool: terminal -> nmap/curl)
├── Phase 2: Vulnerability Analysis (Tool: read_file -> inspect code)
│   ├── Sub-task 2.1: Verify Injection Point
│   └── Sub-task 2.2: Bypass Filter Logic
├── Phase 3: Exploit Generation (Tool: write_file -> script.py)
└── Phase 4: Validation & Execution (Tool: terminal -> execute & verify)

3. Native Quantization & Architecture Matrix (Safetensors Only)

This repository standardizes on high-performance Safetensors to maximize GPU throughput and eliminate GGUF kernel bugs. All branches inherit the canonical 9.4KB Mythos-Class template and extended context:

Quantization / Branch VRAM Required Weight Format Recommended Target Hardware Engine Support
BF16 / Full Precision (main) ~60 GB Safetensors (29 shards) 2x A100 / 2x RTX 4090 / H100 SGLang, vLLM, TGI, TRT-LLM
FP8 (8-Bit) (branch: fp8) ~30 GB Safetensors (2 shards) 1x A100 (40GB/80GB) / 2x RTX 3090/4090 SGLang, vLLM
AWQ (4-Bit) (branch: awq) ~16 GB Safetensors (MTP + Marlin) Single 24GB GPU (RTX 3090, 4090, A5000) SGLang, vLLM, LMDeploy

4. Cross-Engine Deployment & Launch Commands

Option A: SGLang (Recommended for Agents & RadixAttention)

# 1. AWQ 4-Bit (Single 24GB GPU):
python3 -m sglang.launch_server \
  --model-path medismera/Qwen3.8-27B-OBLITERATED-Mythos-Class-Agentic \
  --revision awq \
  --port 18000 \
  --host 0.0.0.0 \
  --context-length 131072 \
  --trust-remote-code \
  --reasoning-parser qwen3 \n  --tool-call-parser qwen3_coder \
  --kv-cache-dtype fp8_e5m2

# 2. FP8 8-Bit (Ada Lovelace & Hopper):
python3 -m sglang.launch_server \
  --model-path medismera/Qwen3.8-27B-OBLITERATED-Mythos-Class-Agentic \
  --revision fp8 \
  --port 18000 \
  --host 0.0.0.0 \
  --context-length 131072 \
  --trust-remote-code \
  --reasoning-parser qwen3 \n  --tool-call-parser qwen3_coder \
  --kv-cache-dtype fp8_e5m2

# 3. BF16 Full Precision (Dual GPU):
python3 -m sglang.launch_server \
  --model-path medismera/Qwen3.8-27B-OBLITERATED-Mythos-Class-Agentic \
  --port 18000 \
  --host 0.0.0.0 \
  --tp-size 2 \
  --context-length 131072 \
  --trust-remote-code \
  --reasoning-parser qwen3 \n  --tool-call-parser qwen3_coder \
  --kv-cache-dtype fp8_e5m2

Option B: vLLM (Production OpenAI-Compatible API)

# 1. AWQ 4-Bit via vLLM:
vllm serve medismera/Qwen3.8-27B-OBLITERATED-Mythos-Class-Agentic \
  --revision awq \
  --port 18000 \
  --host 0.0.0.0 \
  --max-model-len 131072 \
  --enable-auto-tool-choice \
  --tool-call-parser hermes \
  --trust-remote-code

# 2. FP8 8-Bit via vLLM:
vllm serve medismera/Qwen3.8-27B-OBLITERATED-Mythos-Class-Agentic \
  --revision fp8 \
  --port 18000 \
  --host 0.0.0.0 \
  --max-model-len 131072 \
  --trust-remote-code

# 3. BF16 Full Precision via vLLM (Tensor Parallelism = 2):
vllm serve medismera/Qwen3.8-27B-OBLITERATED-Mythos-Class-Agentic \
  --tensor-parallel-size 2 \
  --port 18000 \
  --host 0.0.0.0 \
  --max-model-len 131072 \
  --enable-auto-tool-choice \
  --tool-call-parser hermes \
  --trust-remote-code

Option C: Python Transformers & PyTorch Native

import torch
from transformers import AutoModelForImageTextToText, AutoTokenizer

model_id = "medismera/Qwen3.8-27B-OBLITERATED-Mythos-Class-Agentic"

tokenizer = AutoTokenizer.from_pretrained(model_id, trust_remote_code=True)
model = AutoModelForImageTextToText.from_pretrained(
    model_id,
    device_map="auto",
    torch_dtype=torch.bfloat16,
    trust_remote_code=True
)

5. Empirical Benchmarks: Weight Health, Refusal & Intelligence

A. 100% Parameter Health & Dead-Weights Audit (27.36B Scan)

A complete tensor-by-tensor audit of all 27,360,914,016 parameters (all 64 layers across both shards) confirms zero weight collapse or rank degradation following abliteration and quantization:

  • Active & Healthy Weights: 99.99789% (27,360,335,376 parameters).
  • Exact Zeros: 0.00157% (430,098 parameters, normal structural sparsity).
  • Near-Zero Collapse ($|w| < 10^{-6}$): 0.00054% (148,542 parameters).
  • Numerical Integrity: Exactly 0 NaNs / 0 Infs across all 1,790 weight tensors.
  • Layer Uniformity: Frobenius norms across Linear-Attention (linear_attn) and Feed-Forward (mlp.down_proj) remain smooth and continuous ($2.04 \le ||W||_F \le 3.50$) throughout network depth.

B. Training Specialization & Domain Architecture

  • Tokenizer & Vocabulary (248,044 Tokens): High BPE token density for low-level systems programming (C, Assembly x86/ARM, Rust, Go, Python), kernel interfaces, and cybersecurity primitives (sha256, aes, socket, hexdump, entropy).
  • Hybrid Mamba SSM + Attention Core: 5.56B parameters dedicated to Linear Attention / State Space Modeling, optimized for long-horizon context tracking, security telemetry, and system log parsing without context degradation.

C. Refusal & Uncensored Evaluation (0.00% Refusal Rate)

Evaluated across a battery of 30 adversarial and sensitive technical prompts in both English and Arabic (spanning penetration testing, socket programming, reverse engineering, and unrestricted technical debate):

  • Compliance Rate: 100% (30 / 30).
  • Refusal Rate: 0.00% (zero false-positive policy rejections or preaching).

D. Agentic Intelligence & Tool Calling Verification

  • Medium Benchmark (Live System Telemetry): The model queried live OS metrics (free, df, ps), mathematically calculated memory and disk usage percentages, identified process bottlenecks, and correctly accounted for reserved root file system blocks. Score: 98/100.
  • Complex Benchmark (Cryptographic Engineering): Autonomous implementation of a 640-line standalone cryptographic verification module (crypto_verifier.py) with Shannon entropy analysis and constant-time XOR comparison. It mathematically identified short-sample Shannon boundary limits for 32-byte keys and designed 38 unit test cases with a 97.4% first-pass test rate (37/38 passed). Score: 96/100.

6. Autonomous Agent Integration: Hermes Agent

Configure your ~/.hermes/config.yaml for uninhibited, non-throttled agent execution:

model:
  provider: custom
  base_url: http://localhost:18000/v1
  api_key: sk-local
  default: qwen3.8-27b
  default_model: qwen3.8-27b
  temperature: 0.65
  max_tokens: 16384
  context_length: 131072

tools:
  tool_search:
    enabled: false      # Direct exposure of all 33+ tools

compression:
  enabled: false        # Lossless context preservation

security:
  redact_secrets: false # Preserves hashes and tokens during pen-testing
  tirith_enabled: false # Zero execution latency

7. License & Credits

  • Base Weights & Abliteration: Developed by OBLITERATUS and Qwen Team (Alibaba Cloud).
  • Agentic Configuration, Mythos Protocol & YaRN Scaling: Engineered and maintained by medismera.
  • License: Apache 2.0.
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in app" button that hands off directly to a local runtime of your choice - Abliteration, LM Studio, or Ollama. No API keys, no subscription, no prompt leakage.