← back to catalog · registered 2026-08-22 13:56

acidsound/LFM2.5-2.6B-Uncensored-ONNX

acidsound Lfm 2.6B second-order
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/acidsound%2FLFM2.5-2.6B-Uncensored-ONNX"
Response includes
  • classification m-uncensored
  • files 8
  • hub_downloads_all_time 196
  • author_summary 2 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
196
34 last 30d - stable
Likes
0
Model age
2mo ago
created 2026-08-07
Downloads over time
Now211→from119↑77%
114150185220119 on Aug 5211 on Oct 11211 on Oct 8AugSepOct
Aug 5 → Oct 11 · 50 snapshots · spans 67 days

Genealogy 0 direct forks

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Variants by this author 2 formats · 184 downloads combined

The same weights this author released in different packaging. Pick the format that matches your runtime.

Metadata

License
other
Languages
en zh ko ja fr de es
Tags
transformers.js onnx lfm2 text-generation onnxruntime webgpu lfm2.5 uncensored conversational en zh ko

Related

Total size
0 B
Files
8
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2026-08-07 11:32

Files by quantization

Auxiliary files 8 files 17.1 MB
tokenizer.json 17.1 MB 695be780 download
LICENSE 10.3 KB 25e731c6 download
chat_template.jinja 5.32 KB d63583bc download
README.md 3.24 KB e8beb9cd download
config.json 1.66 KB 47e65c57 download
.gitattributes 1.65 KB 491acd84 download
tokenizer_config.json 363 B e3eefcd5 download
generation_config.json 270 B e46912c1 download

README current version from Hugging Face


library_name: transformers.js
license: other
license_name: lfm1.0
license_link: LICENSE
pipeline_tag: text-generation
language:

  • en
  • zh
  • ko
  • ja
  • fr
  • de
  • es
    tags:
  • onnx
  • onnxruntime
  • webgpu
  • lfm2.5
  • uncensored
  • text-generation
    base_model:
  • SC117/LFM2.5-2.6B-Uncensored
    base_model_relation: finetune

LFM2.5-2.6B-Uncensored ONNX

ONNX Runtime / Transformers.js exports of SC117/LFM2.5-2.6B-Uncensored. These files are intended for browser and cross-platform inference, especially WebGPU.

This repository contains the two web-oriented variants:

Variant File data size Recommended use
Q4 1.85 GB Smaller download; WebGPU/CPU/GPU
Q4F16 1.67 GB Default WebGPU choice; FP16 runtime and KV cache

Q4F16 is not a GGUF quantization. It uses INT4 weights with FP16 runtime tensors and KV cache. It is therefore not interchangeable with Q4_K_M.

Transformers.js

npm install @huggingface/transformers
import { pipeline, TextStreamer } from "@huggingface/transformers";

const generator = await pipeline(
  "text-generation",
  "acidsound/LFM2.5-2.6B-Uncensored-ONNX",
  {
    device: "webgpu",
    dtype: "q4f16", // use "q4" for the Q4 variant
  },
);

const messages = [
  { role: "system", content: "You write concise English and Chinese image prompts." },
  { role: "user", content: "Create a cinematic prompt for a rainy neon street." },
];

const output = await generator(messages, {
  max_new_tokens: 256,
  temperature: 0.3,
  repetition_penalty: 1.05,
  streamer: new TextStreamer(generator.tokenizer, {
    skip_prompt: true,
    skip_special_tokens: true,
  }),
});

console.log(output[0].generated_text.at(-1).content);

The first browser load downloads roughly 1.7–1.9 GB. Use a progress indicator and let the browser cache the model for subsequent sessions. WebGPU support is required for the recommended Q4F16 path.

Files

config.json
generation_config.json
tokenizer.json
tokenizer_config.json
chat_template.jinja
onnx/model_q4.onnx
onnx/model_q4.onnx_data
onnx/model_q4f16.onnx
onnx/model_q4f16.onnx_data

The transformers.js_config section in config.json maps the external-data files and selects FP32 KV cache for Q4 or FP16 KV cache for Q4F16.

Provenance

  • Source model: SC117/LFM2.5-2.6B-Uncensored
  • Source revision: 578ad81f31061f16797ce9f5451835f59d1ad2d5
  • Exporter: Liquid4All/onnx-export
  • Exporter revision: 9a23ddd23035165f7414a5de3220a51e85780f64
  • Export command: LiquidONNX LFM2 builder, block size 32, symmetric Q4; Q4F16 converted from the Q4 graph with FP16 runtime/cache tensors.

SHA-256

File SHA-256
onnx/model_q4.onnx 7AB036E314110DB418B37A36175C5998E9EF441FC88D637523CD6D2A6FF2DE0D
onnx/model_q4.onnx_data F1EBAACBCDFDC164E415E38E47F480B3505F53D7A9D9407332E5C6FDC647C64C
onnx/model_q4f16.onnx 0E17A86E3E5D0FE6EF08967843D74BD79B40F4FA916ADF1B4572A5C2DF546D8A
onnx/model_q4f16.onnx_data 72E0B2CABD22A13F5553410FCF9D91BA8A6ECAEB99998C71CAF176B4045CB919

License

The source model is distributed under the LFM Open License 1.0. See LICENSE and the source model card for the applicable terms.

README history 1 version

The author's README evolved over time. Click a version to see its content at that point.

  1. 2026-08-07Add files using upload-large-folder tool9e2036d3.2 KB
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration