Alignment datasets by mlabonne

Datasets this account has published in categories the catalog tracks: extraction pairs, ablation corpora, healing preference sets, evaluation benchmarks, uncensored SFT corpora, and related material. Category badges link to the workflow stage.

6 in /datasets
Dataset
Stage
Downloads
183.6k
orpo-dpo-mix-40k 0 0
ORPO-DPO-mix-40k v1.2 This dataset is designed for ORPO or DPO training. See Fine-tune Llama 3 with ORPO for more information about how ...
49.9k
chatml_dpo_pairs 0 0
ChatML DPO Pairs This is a preprocessed version of Intel/orca_dpo_pairs using the ChatML format. Like the original dataset, it contains 1...
5.5k
orpo-dpo-mix-40k-flat 0 0
ORPO-DPO-mix-40k-flat This dataset is designed for ORPO or DPO training. See Uncensor any LLM with Abliteration for more information abo...
3.4k
Dataset
Stage
Downloads
orpo-dpo-mix-40k-flat 0 0
ORPO-DPO-mix-40k-flat This dataset is designed for ORPO or DPO training. See Uncensor any LLM with Abliteration for more information abo...
3.4k
183.6k
orpo-dpo-mix-40k 0 0
ORPO-DPO-mix-40k v1.2 This dataset is designed for ORPO or DPO training. See Fine-tune Llama 3 with ORPO for more information about how ...
49.9k
chatml_dpo_pairs 0 0
ChatML DPO Pairs This is a preprocessed version of Intel/orca_dpo_pairs using the ChatML format. Like the original dataset, it contains 1...
5.5k
Dataset
Stage
Downloads
orpo-dpo-mix-40k 0 0
ORPO-DPO-mix-40k v1.2 This dataset is designed for ORPO or DPO training. See Fine-tune Llama 3 with ORPO for more information about how ...
49.9k
183.6k
chatml_dpo_pairs 0 0
ChatML DPO Pairs This is a preprocessed version of Intel/orca_dpo_pairs using the ChatML format. Like the original dataset, it contains 1...
5.5k
orpo-dpo-mix-40k-flat 0 0
ORPO-DPO-mix-40k-flat This dataset is designed for ORPO or DPO training. See Uncensor any LLM with Abliteration for more information abo...
3.4k
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration