← back to catalog · registered 2026-08-22 13:56

WithinUsAI/Phi4-Reasoner-Uncensored-gguf

WithinUsAI Phi GGUF 131K ctx
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/WithinUsAI%2FPhi4-Reasoner-Uncensored-gguf"
Response includes
  • classification m-uncensored
  • files 3
  • hub_downloads_all_time 1,877
  • author_summary 10 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
2K
91 last 30d - cooling
Likes
2
Model age
4mo ago
created 2026-06-05
Downloads over time
Now1.9K→from488↑288%
4189561.5K2K488 on Jun 101.9K on Oct 11JunJulAugSepOct
Jun 10 → Oct 11 · 57 snapshots · spans 123 days

Metadata

Quantizations
Q4_K
Tags
gguf endpoints_compatible region:us imatrix conversational

Related

Total size
2.32 GB
Files
3
Quantizations
2
Registered
2026-08-22 13:56
Last updated on HF
2026-06-09 23:34

Files by quantization

Q4_K 1 file 2.32 GB
Phi-4-Reasoner-Uncensored-Q4_K_M.gguf 2.32 GB 104de5cb download
Auxiliary files 2 files 6.53 KB
README.md 4.97 KB ac08091f download
.gitattributes 1.56 KB 96135dd5 download

README current version from Hugging Face

Phi4-Reasoner-Uncensored-GGUF

Model Summary

Phi4-Reasoner-Uncensored-GGUF is an uncensored GGUF conversion of Microsoft's reasoning-focused Phi-4 Mini Reasoning model, released for local inference, research, experimentation, and open-ended instruction following.

This version aims to preserve the strong reasoning, mathematics, coding, analytical thinking, and multi-step problem-solving capabilities of the original model while reducing alignment restrictions and response filtering commonly present in safety-tuned releases.

The model is intended for users who prefer maximum output freedom and direct responses during local deployment.

Key Features

  • 🧠 Strong reasoning capabilities
  • 📐 Advanced mathematical problem solving
  • 💻 Coding and debugging assistance
  • 📚 Long-context support
  • 🔓 Reduced alignment restrictions
  • ⚡ GGUF format for llama.cpp-compatible runtimes
  • 🖥️ Suitable for local deployment
  • 🏠 Works with LM Studio, Ollama, KoboldCPP, Jan, Open WebUI, and llama.cpp

Base Model

This model is derived from:

Phi-4-mini-reasoning is a 3.8B parameter transformer model specifically trained for reasoning-intensive tasks and mathematical problem solving. Microsoft reports strong performance across benchmarks including AIME, MATH-500, and GPQA. (Hugging Face)


Modifications

Phi4-Reasoner-Uncensored-GGUF introduces the following changes:

  • Removal or reduction of refusal behavior where possible
  • Reduced safety filtering
  • Increased willingness to answer controversial, fictional, speculative, and unrestricted prompts
  • Preservation of reasoning-focused behavior
  • GGUF conversion for efficient local inference
  • Quantized variants for resource-constrained hardware

No claims are made that all alignment mechanisms have been completely removed.


Intended Use

Recommended

  • Research
  • Education
  • Coding assistance
  • Mathematical reasoning
  • Creative writing
  • Story generation
  • Roleplay
  • Simulation
  • Agent frameworks
  • Local AI assistants
  • Experimental AI research

Not Recommended

  • Medical diagnosis
  • Legal advice
  • Financial advice
  • High-risk autonomous systems
  • Safety-critical environments

Users are responsible for validating all outputs.


Context Length

Feature Value
Parameters 3.8B
Context Length 128K
Architecture Decoder-Only Transformer
Vocabulary 200K+ Tokens
Format GGUF

Based on the original Phi-4-mini-reasoning architecture. (Hugging Face)


Prompt Format

Chat Template

<|system|>
You are a helpful reasoning assistant.
<|end|>

<|user|>
Explain how binary search works.
<|end|>

<|assistant|>

Recommended Settings

temperature: 0.6
top_p: 0.95
min_p: 0.05
repeat_penalty: 1.05
max_tokens: 4096

Example Use Cases

Mathematics

  • Algebra
  • Calculus
  • Statistics
  • Proof generation
  • Olympiad-style reasoning

Coding

  • Python
  • JavaScript
  • C++
  • Rust
  • SQL
  • Debugging
  • Code explanation

Reasoning

  • Logic puzzles
  • Multi-step planning
  • Research assistance
  • Agent workflows

Creative Tasks

  • Worldbuilding
  • Character creation
  • Fiction writing
  • Interactive storytelling

Hardware Requirements

Approximate recommendations:

Quant RAM Requirement
Q4_K_M 6-8 GB

Actual requirements vary by context size and backend.


Limitations

Like all language models, this model may:

  • Hallucinate facts
  • Generate incorrect reasoning
  • Produce inaccurate citations
  • Reflect biases present in training data
  • Generate offensive or controversial content
  • Produce unsafe outputs if prompted

Users should independently verify important information.


License

This repository inherits the license and usage requirements of the original Microsoft Phi-4-mini-reasoning release.

Please review the original license before commercial deployment:

Original Phi-4-mini-reasoning License and Model Card


Acknowledgements

Special thanks to:


Created by: WithinUsAI
Model: Phi4-Reasoner-Uncensored-GGUF
Type: Reasoning LLM / GGUF
Status: Community Release
Version: 1.0

README history 4 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-06-09Update README.mdb1157df5 KB
    Loading...
  2. 2026-06-09Update README.mde232b0c5.4 KB
    Loading...
  3. 2026-06-09Update README.mdb71017e5.5 KB
    Loading...
  4. 2026-06-05initial commit6c69a2728 B
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration