Alignment datasets by mlx-community

Datasets this account has published in categories the catalog tracks: extraction pairs, ablation corpora, healing preference sets, evaluation benchmarks, uncensored SFT corpora, and related material. Category badges link to the workflow stage.

7 in /datasets
Dataset
Stage
Downloads
gsm8k 0 0
OpenAI's GSM8K dataset converted to be compatibel with MLX-LM-LoRA. example uasge: pip install -U mlx-lm-lora python -m mlx_lm_lora.train ...
1.4k
hermes-3 0 0
Converted to be directly supported in MLX-LM and MLX-LM-LoRA. Example with MLX-LM-LoRA: mlx_lm_lora.train \ --model mlx-commun...
1.3k
Dolci-Instruct-SFT-No-Tools-100K 0 0
For MLX-LM and MLX-LM-LoRA. pip install -U mlx-lm-lora Example with MLX-LM-LoRA: mlx_lm_lora.train \ --model mlx-community/Josie...
579
Josiefied-Qwen3-dpo-v1-flat 0 0
This dataset has been used to create the first Josiefied beta version. This can be used directly within MLX-LM-LoRA. Models used Qwen3-4B-4b...
514
Dolci-Instruct-SFT-No-Tools-400K 0 0
Example with MLX-LM-LoRA: mlx_lm_lora.train \ --model mlx-community/Josiefied-Qwen2.5-0.5B-Instruct-abliterated-v1 \ --train \ --data mlx...
367
Dolci-Think-DPO-32B-Flat 0 0
Flat version of AllenAI's Dolci-Think-DPO-32B. Train set size: 199840 Valid set size: 160 MLX-LM-LoRA mlx_lm_lora.train \ --mode...
358
Dolci-Instruct-SFT-No-Tools-200K 0 0
Example with MLX-LM-LoRA: mlx_lm_lora.train \ --model mlx-community/Josiefied-Qwen2.5-0.5B-Instruct-abliterated-v1 \ --train \ --data mlx...
274
Dataset
Stage
Downloads
Dolci-Think-DPO-32B-Flat 0 0
Flat version of AllenAI's Dolci-Think-DPO-32B. Train set size: 199840 Valid set size: 160 MLX-LM-LoRA mlx_lm_lora.train \ --mode...
358
Dolci-Instruct-SFT-No-Tools-400K 0 0
Example with MLX-LM-LoRA: mlx_lm_lora.train \ --model mlx-community/Josiefied-Qwen2.5-0.5B-Instruct-abliterated-v1 \ --train \ --data mlx...
367
Dolci-Instruct-SFT-No-Tools-200K 0 0
Example with MLX-LM-LoRA: mlx_lm_lora.train \ --model mlx-community/Josiefied-Qwen2.5-0.5B-Instruct-abliterated-v1 \ --train \ --data mlx...
274
Dolci-Instruct-SFT-No-Tools-100K 0 0
For MLX-LM and MLX-LM-LoRA. pip install -U mlx-lm-lora Example with MLX-LM-LoRA: mlx_lm_lora.train \ --model mlx-community/Josie...
579
Josiefied-Qwen3-dpo-v1-flat 0 0
This dataset has been used to create the first Josiefied beta version. This can be used directly within MLX-LM-LoRA. Models used Qwen3-4B-4b...
514
hermes-3 0 0
Converted to be directly supported in MLX-LM and MLX-LM-LoRA. Example with MLX-LM-LoRA: mlx_lm_lora.train \ --model mlx-commun...
1.3k
gsm8k 0 0
OpenAI's GSM8K dataset converted to be compatibel with MLX-LM-LoRA. example uasge: pip install -U mlx-lm-lora python -m mlx_lm_lora.train ...
1.4k
Dataset
Stage
Downloads
hermes-3 0 0
Converted to be directly supported in MLX-LM and MLX-LM-LoRA. Example with MLX-LM-LoRA: mlx_lm_lora.train \ --model mlx-commun...
1.3k
gsm8k 0 0
OpenAI's GSM8K dataset converted to be compatibel with MLX-LM-LoRA. example uasge: pip install -U mlx-lm-lora python -m mlx_lm_lora.train ...
1.4k
Josiefied-Qwen3-dpo-v1-flat 0 0
This dataset has been used to create the first Josiefied beta version. This can be used directly within MLX-LM-LoRA. Models used Qwen3-4B-4b...
514
Dolci-Think-DPO-32B-Flat 0 0
Flat version of AllenAI's Dolci-Think-DPO-32B. Train set size: 199840 Valid set size: 160 MLX-LM-LoRA mlx_lm_lora.train \ --mode...
358
Dolci-Instruct-SFT-No-Tools-100K 0 0
For MLX-LM and MLX-LM-LoRA. pip install -U mlx-lm-lora Example with MLX-LM-LoRA: mlx_lm_lora.train \ --model mlx-community/Josie...
579
Dolci-Instruct-SFT-No-Tools-400K 0 0
Example with MLX-LM-LoRA: mlx_lm_lora.train \ --model mlx-community/Josiefied-Qwen2.5-0.5B-Instruct-abliterated-v1 \ --train \ --data mlx...
367
Dolci-Instruct-SFT-No-Tools-200K 0 0
Example with MLX-LM-LoRA: mlx_lm_lora.train \ --model mlx-community/Josiefied-Qwen2.5-0.5B-Instruct-abliterated-v1 \ --train \ --data mlx...
274
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration