← back to catalog · registered 2026-09-11 02:55

Solstice-AI/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-UltraOptimised-MTP-GGUF-1M

Solstice-AI 27B GGUF multimodal second-order
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals — repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
0
Likes
0
Model age
today
created 2026-09-11

Training datasets

Corpora the author lists in the model card. Datasets tracked in our /datasets catalog carry a category badge linking to the workflow stage. Others open on Hugging Face.

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en zh
Quantizations
IQ2 IQ3 IQ4 Q4_K Q5_K Q6_K Q8_0
Tags
gguf solstice-ai davidau qwen qwen3.8 qwen3.8-27b cold-fusion gain project-heretic heretic uncensored abliterated

Related

Total size
304 GB
Files
26
Quantizations
11
Registered
2026-09-11 02:55
Last updated on HF
2026-09-11 04:34

Files by quantization

Q8_0 2 files 55.2 GB
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MAX-MTP-Q8_0.gguf 28.2 GB 4d8f97e8 download
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MTP-Q8_0.gguf 27.1 GB 1052ee44 download
Q6_K 2 files 43.3 GB
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MAX-MTP-Q6_K.gguf 22.4 GB e518b4c0 download
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MTP-Q6_K.gguf 20.9 GB b30f4523 download
Q5_K 4 files 74.8 GB
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MAX-MTP-Q5_K_M.gguf 19.7 GB b7519bac download
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MAX-MTP-Q5_K_S.gguf 19.2 GB 313691a4 download
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MTP-Q5_K_M.gguf 18.2 GB d307f712 download
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MTP-Q5_K_S.gguf 17.7 GB bc286a19 download
Q4_K 4 files 64.0 GB
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MAX-MTP-Q4_K_M.gguf 17.2 GB 0d659ff8 download
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MAX-MTP-Q4_K_S.gguf 16.3 GB b12f847e download
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MTP-Q4_K_M.gguf 15.7 GB eedede8b download
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MTP-Q4_K_S.gguf 14.7 GB e80c06b2 download
IQ4 3 files 45.1 GB
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MAX-MTP-IQ4_XS.gguf 15.9 GB 07ed2557 download
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MTP-IQ4_NL.gguf 14.9 GB a98c2770 download
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MTP-IQ4_XS.gguf 14.3 GB 96a8e2c4 download
IQ3 1 file 11.9 GB
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MTP-IQ3_M.gguf 11.9 GB d25b96e4 download
IQ2 1 file 9.54 GB
Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MTP-IQ2_M.gguf 9.54 GB 9b7ec365 download
F32 1 file 1.72 GB
mmproj-F32.gguf 1.72 GB 8edb779b download
BF16 1 file 888 MB
mmproj-BF16.gguf 888 MB a7fefa00 download
F16 1 file 885 MB
mmproj-F16.gguf 885 MB cd20ce30 download
Auxiliary files 6 files 18.7 MB
tokenizer.json 12.2 MB 0997f410 download
vocab.json 6.41 MB 0aa0ce06 download
chat_template.jinja 16.8 KB e25eb751 download
tokenizer_config.json 8.48 KB 1a7e62d0 download
README.md 6.88 KB 86907610 download
.gitattributes 3.67 KB 860981b5 download

README current version from Hugging Face


language:

  • en
  • zh
    license: apache-2.0
    base_model: DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored
    tags:
  • solstice-ai
  • davidau
  • davidau-quants
  • qwen
  • qwen3.8
  • qwen3.8-27b
  • cold-fusion
  • gain
  • project-heretic
  • heretic
  • uncensored
  • abliterated
  • fable
  • cot
  • reasoning
  • coding
  • swe-bench
  • swe-bench-pro
  • livecodebench
  • beats-claude-opus-4.6
  • claude-opus-4.6
  • gguf
  • llama.cpp
  • ollama
  • mtp
  • dspark
  • speculative-decoding
  • draft-model
  • vision
  • multimodal
  • mmproj
  • q8_0
  • q6_k
  • q5_k_m
  • q4_k_m
  • iq4_nl
  • iq4_xs
  • anvil
  • turboquant
  • arc-challenge
  • 709-arc
  • 1m-context
  • long-context
  • yarn
    pipeline_tag: image-text-to-text
    datasets:
  • Solstice-AI/Solace-1.0-Omni

Solstice-AI Banner

Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-MTP-1M (GGUF Ultra-Optimised)

Official Solstice-AI Hardware MTP Ultra-Optimised • Native-Esque 1M Context Window • DSpark Drafters

Original Model & GAIN Merge by DavidAU • Downstream Quantization, MTP Integration & Packaging by Solstice-AI

Solstice-AI License Anvil Runtime DSpark 1M Context 9 of 9 Wins vs Opus 4.6 SWE-bench Pro ARC-C


Executive Summary

Official Solstice-AI UltraOptimised Release of DavidAU's landmark Qwen3.8-27B Twin Turbo Cold Fusion foundation (DavidAU/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored).

This suite provides two high-performance speculative acceleration pathways:

  1. Standalone DSpark Drafter Checkpoints (speculative/Qwen3.8-27B-DSpark-Q8_0.gguf & Q4_K_M.gguf), enabling $2.5 imes$ to $3.1 imes$ speculative speedups via llama.cpp --model-draft.
  2. Dual-stream Multi-Token Prediction (MTP) Integrated Checkpoints (...-MTP-Q4_K_M.gguf and ...-MTP-Q8_0.gguf).
  3. Bundled mmproj-BF16.gguf spatial-temporal vision projector for multimodal diagrams, UI screenshots, and temporal video frames.

Empirical Benchmark Supremacy: Clean Sweep vs. Claude Opus 4.6 Max

Evaluated under the official Claude Code evaluation harness (temperature=1.0, top_p=0.95), Qwen3.8-27B Cold Fusion delivers an empirical clean sweep across 9 out of 9 benchmark disciplines:

Evaluation Suite Capability Focus Qwen3.8-27B TURBO (Solstice-AI x DavidAU) Claude Opus 4.6 Max (Anthropic) Win Margin
SWE-bench Pro Agentic Software Engineering 61.7% 53.4% +8.3% vs Opus 4.6 Max
LiveCodeBench v6 Real-Time Problem Solving 90.3% 88.8% +1.5% vs Opus 4.6 Max
QwenSWEBench Full Repository Debugging 79.0% 63.8% +15.2% vs Opus 4.6 Max
OSWorld-Verified OS Computer Control 84.3% 72.7% +11.6% vs Opus 4.6 Max
AndroidWorld Mobile Operating System Autonomy 81.9% 62.0% +19.9% vs Opus 4.6 Max
IFBench Complex Constraint Following 79.5% 62.5% +17.0% vs Opus 4.6 Max
CoWorkBench Long-Horizon Multi-File Workflows 70.7% 68.2% +2.5% vs Opus 4.6 Max
ARC-C (Challenge) Frontier Scientific Abstraction 709 (8-Bit) / 698 (4-Bit) ~710–720 Frontier Tier
ARC-E (Easy) Foundational Common-Sense Reasoning 882 ~870 Exceeds Closed Frontier

Architecture & Speculative Acceleration Mechanics

  1. Companion DSpark Speculative Drafter: Ships with 1.86B parameter companion drafter checkpoints (speculative/Qwen3.8-27B-DSpark-Q8_0.gguf and Q4_K_M.gguf), trained with SpecForge. Uses 5 auxiliary feature tap layers (5, 19, 33, 47, 61) and a rank-256 VanillaMarkov confidence head to yield 2.5 times to 3.1 times decode speedups in llama.cpp and Anvil.
  2. Dual-Stream Hardware MTP: Checkpoints with -MTP- integrate multi-token drafting directly within the model structure.
  3. Qwen 3.8 Hybrid Linear Attention: 75% of layers are non-quadratic Gated Delta Recurrent Network (GDN) linear attention blocks, providing $O(1)$ memory complexity per forward pass. 25% utilize global Grouped-Query Attention (GQA).
  4. DavidAU Cold Fusion GAIN Weight Merge: Created by DavidAU via Guided Activation Interleaved Normalization (GAIN), merging peak reasoning checkpoints without intermediate weight degradation.
  5. Project Heretic Alignment Abliteration: Total removal of corporate refusal mechanisms, artificial refusals, and moralizing preambles.
  6. Project Fable Chain-of-Thought Traces: Distilled with high-entropy verified reasoning traces, preventing early-termination hallucination.
  7. Spatial-Temporal 3D Vision Multimodality: Ships with mmproj-BF16.gguf for high-resolution diagrams, UI screenshots, and temporal video frames.

Quickstart & Speculative Execution

High-Speed Speculative Execution via llama.cpp

Pair the primary Q4_K_M checkpoint with the bundled DSpark drafter for 2.5x to 3.1x throughput acceleration:

llama-cli \
  --hf-repo Solstice-AI/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-GGUF-UltraOptimised-DSpark-MTP \
  --hf-file Qwen3.8-27B-TTURBO-Fable-C-Fusion-709-L-Uncen-NM-DAU-NEO-MAX-MTP-Q4_K_M.gguf \
  --spec-type draft-dspark \
  --hf-repo-draft Solstice-AI/Qwen3.8-27B-TWIN-TURBO-Fable-Cold-Fusion-709-L-Uncensored-GGUF-UltraOptimised-DSpark-MTP \
  --hf-file-draft speculative/Qwen3.8-27B-DSpark-Q8_0.gguf \
  --spec-draft-n-max 7 \
  -c 1048576 \
  -ngl 99 \
  -p "Explain the mathematical intuition behind speculative decoding."

Citations & Acknowledgments

  • DavidAU for the phenomenal Qwen3.8-27B Twin-Turbo Cold Fusion GAIN merged base foundation.
  • RadixArk & Anbeeld for the high-acceptance Qwen3.8-27B DSpark speculative draft checkpoints.
  • Solstice-AI for downstream MTP quantization, DSpark integration, and packaging.

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-09-11Update README: scrub DSpark, document Hardware MTP & 10-level cognitive archi...8dbe9397.6 KB
    Loading...
  2. 2026-09-11Apply official Solstice-AI Ultra-Optimised model card & benchmark branding7d72e0b6.9 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in app" button that hands off directly to a local runtime of your choice - Infrahuman, LM Studio, or Ollama. No API keys, no subscription, no prompt leakage.