← back to catalog · registered 2026-08-22 13:56

D4pp3rD1scourse/Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-DFlash

D4pp3rD1scourse Qwen 774M GGUF second-order 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/D4pp3rD1scourse%2FQwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive-DFlash"
Response includes
  • classification m-uncensored
  • files 8
  • hub_downloads_all_time 1,910
  • author_summary 1 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
2K
565 last 30d - stable
Likes
0
Model age
8w ago
created 2026-08-15
Downloads over time
Now2.1K→from885↑134%
8261.3K1.7K2.2K885 on Aug 192.1K on Oct 11AugSepOct
Aug 19 → Oct 11 · 48 snapshots · spans 53 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Quantizations
BF16 Q4_K
Tags
gguf safetensors qwen3 dflash speculative-decoding draft-model qwen3.5 dgx-spark custom_code base_model:HauhauCS/Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive base_model:quantized:HauhauCS/Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive license:apache-2.0

Related

Total size
3.32 GB
Files
8
Quantizations
3
Registered
2026-08-22 13:56
Last updated on HF
2026-08-15 14:29

Files by quantization

BF16 1 file 1.45 GB
epoch-011-BF16.gguf 1.45 GB 1d43c144 download
Q4_K 1 file 441 MB
epoch-011-Q4_K_M.gguf 441 MB 0a730eca download
Auxiliary files 6 files 1.44 GB
model.safetensors 1.44 GB 818edcfc download
training.sanitized.json 7.44 KB da0bdddc download
README.md 3.92 KB a56e2731 download
.gitattributes 1.59 KB c012db3a download
config.json 1.19 KB 20f20e46 download
SHA256SUMS 426 B 75d370a7 download

README current version from Hugging Face


license: apache-2.0
base_model:

  • HauhauCS/Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive
  • z-lab/Qwen3.5-122B-A10B-DFlash
    tags:
  • dflash
  • speculative-decoding
  • draft-model
  • gguf
  • qwen3.5
  • dgx-spark
    library_name: gguf

Qwen3.5 122B HauhauCS Aggressive DFlash draft

This repository contains a trained DFlash draft for the exact target:

HauhauCS/Qwen3.5-122B-A10B-Uncensored-HauhauCS-Aggressive, tested with its Q4_K_M GGUF (approximately 74 GB).

It is not a standalone language model. Pair it with the target in a DFlash-capable speculative-decoding runtime. The target weights are not duplicated here.

Important: Hugging Face may display automatically generated llama.cpp, Ollama, or local-app commands for repositories containing GGUF files. Those generic commands are not sufficient for this draft. Load this repository as the draft model alongside the linked HauhauCS target, using the patched DFlash runtime and launch instructions in the companion GitHub repository.

Release files

  • epoch-011-Q4_K_M.gguf — recommended llama.cpp draft, 462,634,656 bytes.
  • epoch-011-BF16.gguf — BF16 GGUF draft.
  • model.safetensors — framework checkpoint.
  • config.json — DFlash architecture configuration.
  • training.sanitized.json — training metadata with host-specific paths removed.
  • SHA256SUMS — release hashes.

Local benchmark

Configuration Median output rate p50 latency
Baseline 25.58 tok/s 4.19 s
Routed DFlash 44.91 tok/s 2.10 s
  • Routed draft acceptance: 77.78%.
  • 8/8 deterministic output comparisons exactly matched the depth-zero baseline.
  • 40/40 routed soak requests completed with 0 errors and 74.37% draft acceptance.

These results describe a small, deterministic local workload. They are not a guarantee for other prompts, runtimes, hardware, concurrency, or decoding settings.

Runtime requirements

The tested llama.cpp path used --spec-type draft-dflash, --spec-draft-model, per-request speculative.n_max, target context 262,144, two parallel slots, thinking disabled, and deterministic decoding.

The required llama.cpp and Speculators patches, launch/training scripts, exact commit pins, and reproducibility record are staged in the companion GitHub repository.

Attribution

Limitations

  • The draft is specialized for the named HauhauCS Aggressive target and tested quant. Compatibility with related Qwen3.5 checkpoints is unverified.
  • Correctness in speculative decoding still depends on the target verifier and compatible runtime behavior.
  • The release does not endorse or reproduce every claim made by the target model publisher.
  • Users are responsible for evaluating safety and suitability in their deployment context.

License

Apache-2.0, subject to the licenses and attribution of the target model, stock DFlash draft, training framework, and runtime components.

Release pair

This repository contains the trained weights, framework checkpoint, sanitized metadata, and checksum manifest. Companion code, patches, launch scripts, and reproducibility details: D4pp3rD1scourse/qwen35-122b-hauhaucs-dflash. Use the draft only with the linked target and compatible patched runtime.

README history 6 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-15docs: finalize paired release linksa3cf8f13.9 KB
    Loading...
  2. 2026-08-15docs: link companion GitHub release code764c4be3.9 KB
    Loading...
  3. 2026-08-15docs: clarify draft-only runtime usage9442c593.7 KB
    Loading...
  4. 2026-08-15docs: correct staging status and runtime metadatad0216973.3 KB
    Loading...
  5. 2026-08-15docs: add model card and release metadata4e2554d3.2 KB
    Loading...
  6. 2026-08-15initial commit67a8ae728 B
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration