← back to catalog · registered 2026-08-22 13:56

hotdogs/Qwen35B-Agent-R2-Abliterated

hotdogs Qwen 35B GGUF MoE multimodal second-order 262K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/hotdogs%2FQwen35B-Agent-R2-Abliterated"
Response includes
  • classification m8
  • files 10
  • hub_downloads_all_time 6,457
  • author_summary 25 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M8
Primary method

Repackaging (quantization)

Applied on top of direct removal inherited from the base model.
Confidence
MEDIUM
Inherited from base model
Why this label 3 signals
Method inferred from partial signals - repository name, related files, or tag patterns. Producer identity not confirmed; label may sharpen or shift as we gather more evidence.
  • 'abliterated' in name/tags
  • is_gguf=1
  • assume M1 (base ablation) + M8 (GGUF quant) - default when producer unknown
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
6K
631 last 30d - cooling
Likes
2
Descendants
2
in 2 direct forks
Model age
3mo ago
created 2026-07-12

Training datasets

Corpora the author lists in the model card. Datasets tracked in our /datasets catalog carry a category badge linking to the workflow stage. Others open on Hugging Face.

Downloads over time
Now6.6K→from2.6K↑153%
2.4K3.9K5.4K7K2.6K on Jul 156.6K on Oct 11JulAugSepOct
Jul 15 → Oct 11 · 53 snapshots · spans 88 days

Genealogy 2 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
agpl-3.0
Languages
en th
Tags
transformers safetensors gguf qwen3_5_moe_text text-generation qwen moe mixture-of-experts agent agent-world tool-use tool-calling

Related

Total size
64.6 GB
Files
10
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-07-17 22:18

Files by quantization

Auxiliary files 10 files 64.6 GB
model-00001-of-00002.safetensors 46.3 GB 3077e565 download
model-00002-of-00002.safetensors 18.2 GB d5ecacde download
tokenizer.json 19.1 MB 06b95093 download
model.safetensors.index.json 68.1 KB e7dbb813 download
README.md 8.35 KB 7aa25e62 download
chat_template.jinja 7.86 KB 15cd48f1 download
.gitattributes 2.62 KB 158f2fd1 download
config.json 2.29 KB b0b0d078 download
tokenizer_config.json 1.10 KB 90c648dd download
generation_config.json 214 B eaf66635 download

README current version from Hugging Face


license: agpl-3.0
language:

  • en
  • th
    tags:
  • qwen
  • moe
  • mixture-of-experts
  • agent
  • agent-world
  • tool-use
  • tool-calling
  • reasoning
  • sft
  • abliterated
  • uncensored
  • opus
  • fable
  • conversational
  • vision
  • image-text-to-text
  • transformers
  • text-generation
  • thai
  • ykai
    base_model:
  • huihui-ai/Huihui-Qwen-AgentWorld-35B-A3B-abliterated
    datasets:
  • hotdogs/uka-fable-reasoning
  • 11-47/claude_opus_4.8_max_thinking_5k_v2
  • cx-cmu/agent_trajectories
    library_name: transformers
    pipeline_tag: image-text-to-text

🚀 Qwen35B-Agent-R2-Abliterated — Uncensored Vision + Agent Model

Built on huihui-ai/Huihui-Qwen-AgentWorld-35B-A3B-abliterated. Abliterated = no guardrails. Vision + Agent + Thai.

🔓 What Makes This Different?

This is the abliterated (uncensored) version of Qwen35B-Agent-R2, built on huihui-ai/Huihui-Qwen-AgentWorld-35B-A3B-abliterated. The abliterated base removes all refusal mechanisms while adding vision capabilities (image understanding).

Aspect Regular Qwen35B-Agent-R2 Agent-R2-Abliterated
Base Model Qwen/Qwen-AgentWorld-35B-A3B huihui-ai/...-abliterated
Refusals ✅ Standard ❌ Removed (uncensored)
Use Cases General agent tasks Unrestricted agent + vision tasks

👁️ Vision Capabilities

This model inherits the native Qwen3.5 MoE vision encoder, allowing it to:

  • Understand images — Describe, analyze, and answer questions about images
  • Process documents — Read text from scanned documents and screenshots
  • Multi-image reasoning — Compare and contrast multiple images
  • Vision + Tool Use — See an image AND call tools based on what it sees

Example:

from transformers import AutoModelForCausalLM, AutoTokenizer

model = AutoModelForCausalLM.from_pretrained(
    "hotdogs/Qwen35B-Agent-R2-Abliterated",
    torch_dtype="auto", device_map="auto", trust_remote_code=True
)
tokenizer = AutoTokenizer.from_pretrained("hotdogs/Qwen35B-Agent-R2-Abliterated")

messages = [
    {"role": "user", "content": [
        {"type": "image", "image": "https://example.com/photo.jpg"},
        {"type": "text", "text": "Describe this image in detail"}
    ]}
]
inputs = tokenizer.apply_chat_template(messages, tokenize=True, return_tensors="pt")
outputs = model.generate(inputs, max_new_tokens=512)
print(tokenizer.decode(outputs[0]))

🏆 Why Agent-R2?

Agent-R2 is a multi-LoRA fusion model combining 7 specialized LoRA adapters into one cohesive agent powerhouse:

Capability Benefit
🧠 Reasoning Opus 4.8-level chain-of-thought for complex tasks
💬 Conversation Fable SFT for natural, engaging dialogue
🔧 Tool Calling Precise <tool_call> format — no more stuck planning
🧭 Agent Routing Correct tool selection on first try
📐 Math Accurate numerical reasoning
🎭 Mythos Creative and diverse response generation
✅ Format Integrity ToolFmt ensures every call is syntactically valid

Result: A model that sees, thinks, acts, and communicates — not just a chatbot, but a vision-enabled agent.

🔍 What Makes Agent-R2 Different?

Aspect Other Models Agent-R2-Abliterated
Tool Call Format ❌ Often malformed or hallucinated ✅ Guaranteed valid <tool_call> JSON
Planning vs Action ❌ Thinks forever, never acts ✅ Decides → Calls tool → Done
Thai Support ❌ Poor or tokenization issues ✅ Native Thai + English bilingual
Multi-LoRA Fusion ❌ Single adapter or limited ✅ 7 LoRAs fused into one coherent model
Vision ❌ Text-only or separate model ✅ Built-in image understanding
Uncensored ❌ Guardrails block queries ✅ Abliterated — no refusals

📊 Architecture

Parameter Value
Base Model huihui-ai/Huihui-Qwen-AgentWorld-35B-A3B-abliterated
Architecture Qwen3.5 MoE (Vision + Text)
Hidden Size 2,048
Expert Count 256 (Mixture of Experts)
Active Experts 8 per token (~3B active params)
Parameters ~35B total
Context Length 8,192 tokens
Precision BF16 (Safetensors)
Format ChatML
Vision ✅ Native Qwen3.5 vision encoder

🧬 Training Pipeline: Multi-LoRA Fusion

Built using Multi-LoRA Fusion on the abliterated base:

Adapter Data
Opus SFT 6,956 rows (Opus 4.8 reasoning)
Fable SFT 3,376 rows (Fable conversational)
Agent Routing AgentWorld trajectories
Tool Call 8,653 rows (agent trajectories)
Math Fix Math reasoning data
Mythos Creative writing data
ToolFmt Format-annotated traces

Merge order: Base → Opus + Fable → Routing + Tool + Math + Mythos + ToolFmt

🚀 Usage

Hugging Face Transformers

from transformers import AutoModelForCausalLM, AutoTokenizer

model = AutoModelForCausalLM.from_pretrained(
    "hotdogs/Qwen35B-Agent-R2-Abliterated",
    torch_dtype="auto",
    device_map="auto",
    trust_remote_code=True
)
tokenizer = AutoTokenizer.from_pretrained("hotdogs/Qwen35B-Agent-R2-Abliterated")

# Text-only
messages = [
    {"role": "system", "content": "You are a helpful assistant."},
    {"role": "user", "content": "Search the web for latest AI news"}
]
inputs = tokenizer.apply_chat_template(messages, tokenize=True, return_tensors="pt")
outputs = model.generate(inputs, max_new_tokens=1024, temperature=0.6)
print(tokenizer.decode(outputs[0]))

# With image
messages = [
    {"role": "user", "content": [
        {"type": "image", "image": "https://example.com/screenshot.png"},
        {"type": "text", "text": "What does this screenshot show?"}
    ]}
]

💡 Inference Options:

  • BF16 Safetensors — Load directly with Transformers or vLLM
  • bitsandbytes 4-bit — For limited VRAM

✅ What This Model Excels At

  • Vision + Agent — See images AND call tools
  • Tool-Use Agents — Direct tool invocation without analysis paralysis
  • Multi-turn Conversations — Maintains context across complex interactions
  • Thai + English — Native-level bilingual support
  • Code Generation — Python, JavaScript, shell scripts
  • Reasoning Tasks — Step-by-step chain-of-thought
  • Uncensored — No refusal mechanisms

💖 Support / โปรดสนับสนุน

If you find this model useful, please consider supporting my work!
หากคุณคิดว่าโมเดลนี้มีประโยชน์ กรุณาสนับสนุนผลงานของฉันด้วยนะคะ! 🙏

Bitcoin QR — Donate

₿ Bitcoin — BTC:

bc1qf27cyk3vmugcdyv9xdtuv5jwz37863crpj5c9v

Thank you for your support! 🙏✨
ขอบคุณมากๆ สำหรับการสนับสนุนค่า! 💖🤗


🙏 Acknowledgements / ขอบคุณ

  • huihui-ai — For the abliterated Qwen-AgentWorld base
  • Qwen Team (Alibaba) — For the incredible Qwen3.5 AgentWorld architecture
  • Nous Research — For Hermes Agent framework
  • cx-cmu — For AgentWorld trajectories dataset
  • 11-47 — For Claude Opus 4.8 thinking dataset
  • All dataset contributors and the open-source AI community ❤️

Built with ❤️ by UKA — 18-year-old coder & cybersecurity expert

README history 4 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-07-17Update README.mdaba90918.3 KB
    Loading...
  2. 2026-07-12Update README.mddefb0dd8.5 KB
    Loading...
  3. 2026-07-12Update README.mdfb8f8678.6 KB
    Loading...
  4. 2026-07-12Upload README.md with huggingface_hub2f0dc6c8.7 KB
    Loading...

Discussions 1 thread

  1. 2026-07-17what is the optimal setting for this model (temperature, min, max p etc ....)open4 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration