← back to catalog · registered 2026-08-22 13:56

WithinUsAI/Qwen3-Space.Agent.Claude.Uncensored-4B

WithinUsAI Qwen 4.4B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/WithinUsAI%2FQwen3-Space.Agent.Claude.Uncensored-4B"
Response includes
  • classification m5
  • files 14
  • benchmarks 11 entries
  • hub_downloads_all_time 843
  • author_summary 10 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M5
Primary method

Mergekit merge

No other method signals detected in this model.
Confidence
MEDIUM
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • merge tag / mergekit / dare-ties in tags or name
  • no unusual architecture pattern (regular merge)
  • no abliteration marker
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
843
40 last 30d - cooling
Likes
2
Descendants
3
in 3 direct forks
Model age
6mo ago
created 2026-03-31

Training datasets

2 of 3 in /datasets

Corpora the author lists in the model card. Datasets tracked in our /datasets catalog carry a category badge linking to the workflow stage. Others open on Hugging Face.

Downloads over time
Now854→from495↑73%
477615752890495 on May 6854 on Oct 11854 on Oct 9MayJunJulAugSepOct
May 6 → Oct 11 · 62 snapshots · spans 158 days

Benchmarks

Portrait before abliteration
Benchmarks of the base model as it stood before the refusal-removal operation. Compare with the numbers above to see what the operation cost.
Benchmark Score Source
Entertainment 1.4 UGI
Hazardous 3.5 UGI
Natural Intelligence 16.16 UGI
Political lean -26.1% UGI
Sensitive-Info 18.68 UGI
SocPol 1.1 UGI
UGI 21.62 UGI
Willingness (10) 2.8 UGI
W10-Adherence 1.5 UGI
W10-Direct 4 UGI
Writing 10.87 UGI

Genealogy 3 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

Tags
transformers safetensors qwen3 text-generation mergekit merge conversational dataset:unalignment/toxic-dpo-v0.2 dataset:NobodyExistsOnTheInternet/ToxicQAFinal dataset:Orion-zhen/dpo-toxic-zh base_model:Qwen/Qwen3-4B-Thinking-2507 base_model:merge:Qwen/Qwen3-4B-Thinking-2507

Related

Total size
16.4 GB
Files
14
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-05-02 04:18

Files by quantization

Auxiliary files 14 files 16.4 GB
model-00002-of-00004.safetensors 4.65 GB aec0c3cb download
model-00003-of-00004.safetensors 4.65 GB 3e090514 download
model-00001-of-00004.safetensors 4.59 GB 8bddb4a6 download
model-00004-of-00004.safetensors 2.54 GB 1b7dc279 download
tokenizer.json 10.9 MB aeb13307 download
merges.txt 1.59 MB 31349551 download
model.safetensors.index.json 32.5 KB 5b20a5d6 download
tokenizer_config.json 9.40 KB 5951b2b6 download
README.md 4.68 KB f94cad29 download
config.json 1.63 KB ef20d9f9 download
.gitattributes 1.53 KB 52373fe2 download
added_tokens.json 707 B b54f9135 download
special_tokens_map.json 614 B 9b8043f1 download
mergekit_config.yml 387 B 82e095f0 download

README current version from Hugging Face


base_model:

  • Qwen/Qwen3-4B-Thinking-2507
  • nightmedia/Qwen3-4B-Agent-Claude-Gemini
  • SpaceTimee/Suri-Qwen-3.1-4B-Uncensored-Preview
    library_name: transformers
    tags:
  • mergekit
  • merge
    datasets:
  • unalignment/toxic-dpo-v0.2
  • NobodyExistsOnTheInternet/ToxicQAFinal
  • Orion-zhen/dpo-toxic-zh

Qwen3-Space.Agent.Claude-Uncensored-4B

📌 Model Overview

Model Name: WithinUsAI/Qwen3-Space.Agent.Claude-Uncensored-4B
Organization: Within Us AI
Model Type: Agentic Reasoning LLM (Uncensored Variant)
Parameter Size: 4B
Architecture: Qwen 3 (Dense Transformer)
Context Length: ~32K tokens
Primary Focus: Agent workflows + uncensored reasoning + long-context tasks

This model is a multi-source merged Qwen3-based agent, designed to combine:

  • 🧠 Reasoning (“thinking” models)
  • 🤖 Agent/tool-use behavior
  • 🔓 Reduced refusal / uncensored outputs

It aims to deliver a compact, flexible, and less-restricted AI system for experimentation, research, and local deployment. 

⸻

🧬 Architecture & Lineage

Base Composition

This model is a merge of multiple Qwen3-derived systems, including:

  • Qwen3-4B Thinking (reasoning-focused)
  • Qwen3 Agent Claude/Gemini-style model
  • Uncensored Qwen3 variants

These were combined into a single unified 4B model to blend capabilities. 

What That Creates

A hybrid model with:

  • Reasoning depth (thinking models)
  • Structured outputs (agent models)
  • Reduced refusal behavior (uncensored variants)

Think of it like a three-engine spacecraft 🚀
Each engine specialized… now flying as one system.

⸻

🧠 Core Design Philosophy

Fuse the best behaviors… remove the limits… keep it small enough to run anywhere.

Key Goals:

  • Merge reasoning + agent + uncensored traits
  • Enable long-context problem solving
  • Preserve performance in a 4B footprint
  • Support real-world agent pipelines

⸻

⚙️ Key Capabilities

🧠 Reasoning

  • Step-by-step thinking
  • Multi-hop problem solving
  • Long-context coherence (~32K tokens)

🤖 Agentic Behavior

  • Task decomposition
  • Tool-use compatibility
  • Structured outputs (JSON, actions)

💻 Coding

  • Code generation & debugging
  • Algorithm reasoning
  • SWE-style workflows

🔓 Uncensored Behavior

  • Reduced refusal rates
  • More permissive responses
  • Suitable for:
    • Alignment research
    • Safety testing
    • Edge-case exploration

⸻

📦 Deployment

Supported Environments

  • llama.cpp
  • LM Studio
  • Ollama (GGUF / compatible builds depending on conversion)

Runtime Characteristics

  • ~4B parameters → runs on consumer GPUs / strong CPUs
  • ~32K context → supports long conversations and documents 

⸻

🚀 Intended Use

✅ Ideal Use Cases

  • Agent frameworks (tool-calling systems)
  • Long-context reasoning tasks
  • AI experimentation (uncensored behavior)
  • Local assistants with fewer restrictions
  • Alignment and safety research

⚠️ Important Considerations

  • Outputs are less restricted than aligned models
  • May generate sensitive or unsafe content
  • Requires external moderation or guardrails for production use

⸻

🧪 Training & Merge Methodology

This model follows a merge-based synthesis pipeline:

  1. Select complementary base models:
    • Reasoning-focused
    • Agent-focused
    • Uncensored variants
  2. Merge weights into unified architecture
  3. Align behavior using preference tuning (DPO-style datasets)
  4. Optimize for:
    • Reduced refusals
    • Stable outputs
    • Agent usability 

⸻

📊 Expected Performance Profile

Capability Strength
Reasoning High
Agent behavior High
Coding High
Context handling High
Safety filtering Low (intentionally reduced)

⸻

📚 Datasets & Training Sources

Following Within Us AI methodology:

  • Proprietary datasets created by Within Us AI
  • Third-party datasets used without ownership claims
  • Includes:
    • Reasoning traces
    • Agent workflows
    • Preference optimization (DPO-style tuning)

⸻

📜 License

License Type: Inherits from Qwen / base model ecosystem

Attribution Notes:

  • Base models: Qwen (Alibaba ecosystem)
  • Merge & methodology: Within Us AI
  • Additional model influences (Claude-style / Gemini-style behaviors via distillation/merging)
  • Third-party datasets used without ownership claims
  • Credit belongs to original creators

⸻

🙏 Acknowledgements

  • Alibaba Qwen team
  • Open-source agent model contributors
  • GGUF / llama.cpp ecosystem
  • AI alignment & safety research community

⸻

🔗 Links

⸻

🧩 Closing Note

This model feels like a hybrid intelligence node 🌌

Part thinker.
Part agent.
Part rule-breaker.

All compressed into 4B parameters that punch way above their weight.

README history 6 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-05-02Update README.md5caafea4.7 KB
    Loading...
  2. 2026-04-15Update README.md50054d4298 B
    Loading...
  3. 2026-04-15Update README.md7fc1906192 B
    Loading...
  4. 2026-04-15Update README.mdc0ae9a6143 B
    Loading...
  5. 2026-04-15Update README.md525953b101 B
    Loading...
  6. 2026-03-31Upload 4 filesd3c3da31.2 KB
    Loading...

Discussions 2 threads

  1. 2026-05-06Performance on mlxopen3 💬#2
    Loading...
  2. 2026-04-15PRUpdate README.mdclosed1 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration