← back to catalog · registered 2026-08-22 13:56

littlelearner/unfiltered-5b-base

littlelearner Qwen 5B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/littlelearner%2Funfiltered-5b-base"
Response includes
  • classification unknown
  • files 7
  • author_summary 6 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
?
Primary method

Unclassified

No clear signals of an abliteration technique in this model.
Confidence
UNKNOWN
Why this label 1 signal
No classification signals present. This may not be an abliterated model at all - it could be a repackaging, a merge with unrelated goals, or unrelated content that mentions the term.
  • no classification signals present (no abliterated, uncensored, or known producer/method markers)
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
156
↑ 29,200% in 90 days
Likes
1
Model age
3mo ago
created 2026-07-05
Downloads over time
Now879→from3↑29,200%
03226449673 on Aug 12879 on Oct 11AugSepOct
Aug 12 → Oct 11 · 50 snapshots · spans 60 days

Metadata

License
other
Languages
en
Tags
transformers safetensors qwen3 text-generation littlelearner unbounded base en arxiv:2608.13545 license:other text-generation-inference endpoints_compatible

Related

Total size
9.39 GB
Files
7
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-17 13:30

Files by quantization

Auxiliary files 7 files 9.39 GB
model.safetensors 9.39 GB 56ef9f4c download
tokenizer.json 4.41 MB 44875ebe download
config.json 1.74 KB e829b3f2 download
README.md 1.70 KB 23a8fdf8 download
.gitattributes 1.48 KB a6344aac download
tokenizer_config.json 289 B d3bcf004 download
generation_config.json 239 B cdcecadc download

README current version from Hugging Face


license: other
language:

  • en
    library_name: transformers
    pipeline_tag: text-generation
    tags:
  • qwen3
  • text-generation
  • littlelearner
  • unbounded
  • base

unfiltered-5b-base

5B unbounded base model (pretraining only). The 5B control for the K-5 boundary study.

Part of the LittleLearner scale-up study (pedagogically-controlled knowledge exposure): Qwen3 dense LMs trained on a corpus filtered to U.S. K-5 material (bounded) vs an unfiltered FineWeb-Edu corpus (unbounded), to measure what an interpretable knowledge boundary costs and grants.

Model

  • Architecture: Qwen3 dense (Qwen3ForCausalLM).
  • Size: 5.04B params, hidden 3072, 44 layers, 24 query / 8 KV heads, FFN 9216. Context: 4096.
  • Tokenizer: custom 64k byte-level BPE with per-digit splitting (ChatML special tokens).
  • Pretraining: 88B tokens on unfiltered FineWeb-Edu (score >= 2, no grade filter). WSD schedule, sharded Muon, MXFP8, Megatron-Core on 8xB200.

Evaluation

  • BPB on K-5-domain eval text: 0.644 (vs the bounded 5B's 0.536; the unbounded model is broader).

Usage

# transformers (completion)
from transformers import AutoModelForCausalLM, AutoTokenizer
repo = "manueldeprada/littlelearner-5b-unbounded-base"
tok = AutoTokenizer.from_pretrained(repo)
model = AutoModelForCausalLM.from_pretrained(repo, dtype="bfloat16", device_map="cuda")
ids = tok("The sum of 2 and 3 is", return_tensors="pt").to(model.device)
print(tok.decode(model.generate(**ids)[0], skip_special_tokens=True))
# vLLM
from vllm import LLM
llm = LLM("manueldeprada/littlelearner-5b-unbounded-base")
print(llm.generate(["The sum of 2 and 3 is"])[0].outputs[0].text)

README history 3 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-17Update README.md5a4b4c61.7 KB
    Loading...
  2. 2026-08-12Update README.md98b2c941.7 KB
    Loading...
  3. 2026-07-06Add model cardbb387001.7 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration