← back to catalog · registered 2026-08-22 13:56

trollek/danube2-1.8b-WizardLM-Evol-V2-Unfiltered

trollek Mistral 1.8B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/trollek%2Fdanube2-1.8b-WizardLM-Evol-V2-Unfiltered"
Response includes
  • classification unknown
  • files 10
  • author_summary 2 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
?
Primary method

Unclassified

No clear signals of an abliteration technique in this model.
Confidence
UNKNOWN
Why this label 1 signal
No classification signals present. This may not be an abliterated model at all - it could be a repackaging, a merge with unrelated goals, or unrelated content that mentions the term.
  • no classification signals present (no abliterated, uncensored, or known producer/method markers)
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · 30-day
26
↑ 1,659% in 90 days
Likes
0
Descendants
3
in 3 direct forks
Model age
2.3y ago
created 2024-07-06

Training datasets

Corpora the author lists in the model card. Datasets tracked in our /datasets catalog carry a category badge linking to the workflow stage. Others open on Hugging Face.

Downloads over time
Now475→from27↑1,659%
017434852227 on Jul 24, 2024475 on Oct 11Jul '24Nov '24Mar '25Jul '25Nov '25MarJul
Jul 24, 2024 → Oct 11 · 155 snapshots · spans 809 days

Genealogy 3 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

License
apache-2.0
Languages
en
Tags
transformers safetensors mistral text-generation llama-factory unsloth conversational en dataset:cognitivecomputations/WizardLM_evol_instruct_V2_196k_unfiltered_merged_split arxiv:2404.02827 base_model:h2oai/h2o-danube2-1.8b-base base_model:finetune:h2oai/h2o-danube2-1.8b-base

Related

Total size
3.41 GB
Files
10
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2024-07-09 09:47

Files by quantization

Auxiliary files 10 files 3.41 GB
model.safetensors 3.41 GB 1dd267f3 download
tokenizer.json 1.71 MB 40974cf3 download
tokenizer.model 482 KB dadfd56d download
README.md 2.48 KB 652f716a download
tokenizer_config.json 1.99 KB 8e806517 download
.gitattributes 1.48 KB a6344aac download
special_tokens_map.json 885 B 688698eb download
config.json 670 B 823fde55 download
generation_config.json 140 B 53d61feb download
added_tokens.json 51.0 B 281abec6 download

README current version from Hugging Face


license: apache-2.0
datasets:

  • cognitivecomputations/WizardLM_evol_instruct_V2_196k_unfiltered_merged_split
    language:
  • en
    library_name: transformers
    base_model: h2oai/h2o-danube2-1.8b-base
    tags:
  • llama-factory
  • unsloth

h2o-danube2 with ChatML template

This model was first fine-tuned with BAdam on cognitivecomputations/WizardLM_evol_instruct_V2_196k_unfiltered_merged_split using LLama-Factory.

Quants

Thanks to mradermacher!

Template

<|im_start|>system
You are a helpful assistant that gives long and detailed answers.<|im_end|>
<|im_start|>user
{{instruction}}<|im_end|>
<|im_start|>assistant
{{response}}<|im_end|>

BAdam config

### model
model_name_or_path: danube2-base-chatml

### method
stage: sft
do_train: true
finetuning_type: full
use_badam: true
badam_switch_mode: ascending
badam_switch_interval: 50
badam_verbose: 1
badam_start_block: 6
seed: 720

### dataset
dataset: wizardlm_evol_v2_196k_unfiltered
template: ninja_chatml
cutoff_len: 8192
overwrite_cache: false
preprocessing_num_workers: 12

### output
output_dir: wizardlm-evol-v2-chatml-badam
logging_steps: 5
save_steps: 1
save_strategy: epoch
plot_loss: true
overwrite_output_dir: false

### train
per_device_train_batch_size: 2
gradient_accumulation_steps: 8
learning_rate: 0.00001
num_train_epochs: 1
lr_scheduler_type: constant_with_warmup
warmup_ratio: 0.01
pure_bf16: true
flash_attn: fa2

### eval
val_size: 0.01
per_device_eval_batch_size: 1
eval_strategy: steps
eval_steps: 1000

BAdam training results

Training Loss Epoch Step Validation Loss
0.6195 0.1050 1000 0.7363
0.6788 0.2100 2000 0.7252
0.689 0.3150 3000 0.7172
0.6707 0.4200 4000 0.7133
0.6674 0.5250 5000 0.7091
0.7365 0.6301 6000 0.7085
0.7037 0.7351 7000 0.7066
0.709 0.8401 8000 0.7041
0.6652 0.9451 9000 0.7042

README history 4 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2024-07-09Update README.mde22bfbe2.5 KB
    Loading...
  2. 2024-07-08Update README.mda6f71472.3 KB
    Loading...
  3. 2024-07-08Update README.md57d10d42.3 KB
    Loading...
  4. 2024-07-06initial commit6a2f47828 B
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration