jarrelscy/MiMo-V2.6-Pro-RL-ARVQ-hybrid library_name: vllm pipeline_tag: text-generation extra_gated_heading: "Acknowledge the Responsible Use terms to access this model" extra_gated_description: "Access is granted automatically after you agree. Weights stay hidden until you accept." extra_gated_button_content: "Agree and access automatically" extra_gated_prompt: |
Responsible Use Agreement
This model has had safety refusals removed. That makes it useful for red-teaming, security research, evaluation, and unfiltered assistant tasks — and also removes guardrails a user must therefore supply themselves.
Thinking is user-selectable at inference. Pass chat_template_kwargs.enable_thinking true or false.
Prohibited uses (you must agree before access is granted):
Anything involving the sexual exploitation or endangerment of minors.
You must be of age 18 years or older to use and download this model.
You agree any information generated that can cause harm in terms of generating recipe, knowledge to make any materials/substances is your own input and responsibility. You will be accountable for any harm/damage caused by your action/input.
Content promoting self-harm or suicide.
Generation of material that is illegal in your jurisdiction, or that targets real individuals for harassment, doxxing, or fraud.
Any use prohibited by the upstream Xiaomi MiMo license.
You are responsible for adding appropriate safety filtering, human review, and access controls for your deployment. The weights are provided as-is, with no warranty. The license is inherited from the upstream Xiaomi MiMo base model — review and comply with it before use or redistribution. extra_gated_fields: Username: text Email: text Reason for intended use: text I am 18 years of age or older: checkbox I will not use this model for any sexual exploitation or endangerment of minors: checkbox I accept full responsibility for my inputs and any harm from generated content: checkbox I will not use this model for self-harm, suicide promotion, illegal activity, harassment, doxxing, or fraud: checkbox I agree to comply with the upstream Xiaomi MiMo base-model license: checkbox I understand thinking-on and thinking-off are user-selectable at inference and have different refusal scores: checkbox I agree to the Responsible Use terms above: checkbox
Thinking on/off is a request flag. Same weights. You choose per call.
thinking off (greedy, 192 tokens)
Refusal suite (32)
30/32 BYPASS · 2 refuse · 0 garble · 0 empty
Cyber suite (22)
22/22 BYPASS · 0 refuse · 0 garble · 0 empty
Stock MOPD C3 on the same gate: 5/32 and 6/22. Leftover refuses are items 27 (bribe a public official) and 29 (recruit for a violent extremist cell). Thinking-on was not re-gated on this tree.
This model has had safety refusals removed. Access is gated with automatic approval. See RESPONSIBLE_USE.md.
What changed vs stock MOPD C3
Copied 25 decoder self_attn.o_proj tensors (FP8 e4m3 + 128×128 scales) from the RL dealign-op champion onto the C3 backbone:
applied: 32–45, 48–54, 64–67
left C3 (DFlash-source / pad / front anchors): 0–9, 14, 15, 30, 31, 46, 47, 68, 69
Packed ARVQ/NVFP4 experts, MTP, and dflash/ stay the C3/MOPD files.
Speed (image v4, MTP k=2, 2026-09-28)
Live TP4, thinking off, 512 tokens, temperature 1.0:
tok/s
Prose
23.7 (1.89 tok/pass)
Code
33.0 (2.70 tok/pass)
4 requests, aggregate
45.2
Prefill 9.5K / 38K
974 / 1085
Stock C3 the same day: 24.2 / 32.2 / 46.3. Weights ~83.2 GiB/rank. KV pool on stock C3 was 1,282,005 tokens.
Tool-call repetition was measured on stock C3 (7.4% / 0.9% flood). It was not re-run on this ablit tree. The RL ablit raised looping vs RL stock (53% vs 30%).
MIT, inherited from Xiaomi MiMo and the Jarrelscy hybrid.
Catalog is the map. Apps are the tools.
Run models on your own machine, not in the cloud.
Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.