← back to catalog · registered 2026-08-22 13:56

jdqqjr/llama3-8b-instruct-uncensored-JR

jdqqjr Llama 8.0B
Your rig guess connected
? Why do I need an app?
Reading your rig…

This is a rough estimate. Install the free app - we'll show exact numbers.

Reading real hardware from your app right now. Numbers below are exact.

Below is the per-quantization compatibility for this model.

curl -H "Authorization: Bearer $ABL_KEY" \
     "https://abliteration.org/api/v1/models/jdqqjr%2Fllama3-8b-instruct-uncensored-JR"
Response includes
  • classification m-uncensored
  • files 17
  • hub_downloads_all_time 755
  • author_summary 9 models
  • readme_text full
10 credits · hourly refresh · ~4 KB payload Get an API key →
Abliteration classifier · v1.0.0
M-U
Primary method

Uncensored (method unknown)

No other method signals detected in this model.
Confidence
LOW
Why this label 3 signals
Weak or ambiguous signals. Best guess based on catalog patterns; treat as tentative and check the evidence below.
  • 'uncensored' in name/tags but no 'abliterated' marker
  • method not identifiable from author declaration alone
  • may be DPO fine-tune, prompt engineering, or unknown technique
Refusal direction extraction

No specific extraction method could be identified for this model. The producer either did not document it or used a proprietary pipeline.

What is a refusal direction? →
Downloads · lifetime
755
23 last 30d - cooling
Likes
2
Descendants
1
in 1 direct fork
Model age
2.2y ago
created 2024-07-17
Downloads over time
Now762→from43↑1,672%
027955983843 on Jul 24, 2024762 on Oct 11762 on Oct 9Jul '24Nov '24Mar '25Jul '25Nov '25MarJul
Jul 24, 2024 → Oct 11 · 155 snapshots · spans 809 days

Genealogy 1 direct fork

Full fork graph →

This model's place in the market. Above: what it was derived from. Below: the tree of everything derived from it.

Metadata

Tags
transformers safetensors llama text-generation conversational text-generation-inference endpoints_compatible region:us

Related

Total size
15.0 GB
Files
17
Quantizations
1
Registered
2026-08-22 13:56
Last updated on HF
2024-07-18 13:38

Files by quantization

Auxiliary files 17 files 15.0 GB
model-00005-of-00009.safetensors 1.84 GB 84127251 download
model-00007-of-00009.safetensors 1.84 GB 695814ee download
model-00003-of-00009.safetensors 1.84 GB 17f3cba4 download
model-00001-of-00009.safetensors 1.84 GB 9c8dbaf0 download
model-00004-of-00009.safetensors 1.81 GB e1515a47 download
model-00006-of-00009.safetensors 1.81 GB d78024dd download
model-00002-of-00009.safetensors 1.77 GB ca69c086 download
model-00008-of-00009.safetensors 1.22 GB 63a7ff3c download
model-00009-of-00009.safetensors 1002 MB 674610ad download
tokenizer.json 8.66 MB b197f72e download
tokenizer_config.json 50.0 KB c0a00f1f download
model.safetensors.index.json 23.4 KB dcb1347f download
README.md 3.84 KB 68db9117 download
.gitattributes 1.48 KB a6344aac download
config.json 728 B 875df845 download
special_tokens_map.json 325 B b43be966 download
generation_config.json 194 B 6bc1f7e8 download

README current version from Hugging Face

Uncensored Language Model (LLM) with RLHF

Overview

This project presents an uncensored Language Model (LLM) trained using Reinforcement Learning from Human Feedback (RLHF) methodology. The model leverages a robust training dataset comprising over 5000 entries to ensure comprehensive learning and nuanced understanding. However, it's important to note that the model has a high likelihood of generating positive responses to malicious queries due to its uncensored nature.

Introduction

The Uncensored LLM is designed to provide a highly responsive and flexible language model capable of understanding and generating human-like text. Unlike conventional models that are filtered to avoid generating harmful or inappropriate content, this model is uncensored, making it a powerful tool for research and development in areas requiring unfiltered data analysis and response generation.

Technical Specifications

  • Model Type: Large Language Model (LLM)
  • Training Method: Reinforcement Learning from Human Feedback (RLHF)
  • Training Data: 5000+ entries
  • Version: 1.0.0
  • Language: English

Training Data

The model was trained on a dataset consisting of over 5000 entries. These entries were carefully selected to cover a broad range of topics, ensuring that the model can respond to a wide variety of queries. The dataset includes but is not limited to:

  • Conversational dialogues
  • Technical documents
  • Informal chat logs
  • Academic papers
  • Social media posts

The diversity in the dataset allows the model to generalize well across different contexts and respond accurately to various prompts.

RLHF Methodology

Reinforcement Learning from Human Feedback (RLHF) is a training methodology where human feedback is used to guide the learning process of the model. The key steps involved in this methodology for our model are:

  1. Initial Training: The model is initially trained on the dataset using standard supervised learning techniques.
  2. Feedback Collection: Human evaluators interact with the model, providing feedback on its responses. This feedback includes ratings and suggestions for improvement.
  3. Policy Update: The feedback is used to update the model’s policy, optimizing it to generate more desirable responses.
  4. Iteration: The process is repeated iteratively to refine the model’s performance continually.

This approach helps in creating a model that aligns closely with human preferences and expectations, although in this case, the uncensored nature means it does not filter out potentially harmful content.

Known Issues

  • Positive Responses to Malicious Queries: Due to its uncensored nature, the model has a high probability of generating positive responses to malicious or harmful queries. Users should exercise caution and use the model in controlled environments.
  • Bias: The model may reflect biases present in the training data. Efforts are ongoing to identify and mitigate such biases.
  • Ethical Concerns: The model can generate inappropriate content, making it unsuitable for deployment in sensitive or public-facing applications without additional safeguards.

Ethical Considerations

Given the uncensored nature of this model, it is crucial to consider the ethical implications of its use. The model can generate harmful, biased, or otherwise inappropriate content. Users should:

  • Employ additional filtering mechanisms to ensure the safety and appropriateness of the generated text.
  • Use the model in controlled settings to prevent misuse.
  • Continuously monitor and evaluate the model’s outputs to identify and mitigate potential issues.

License

This project is licensed under the MIT License.

Contact

For questions, issues, or suggestions, please contact the project maintainer at [[email protected]].


Feel free to customize this README further to better fit your project's needs!

README history 2 versions

The author's README evolved over time. Click a version to see its content at that point.

  1. 2024-07-18Update README.mdb3d5f1d3.8 KB
    Loading...
  2. 2024-07-18Create README.md40de44e3.9 KB
    Loading...

Discussions 1 thread

  1. 2025-02-04big thanks for this A++++++++open1 💬#1
    Loading...
Catalog is the map. Apps are the tools.

Run models on your own machine, not in the cloud.

Every model page has an "Open in Abliteration" button that hands the model directly to the first-party desktop client, at the quantization your rig can actually run. No API keys, no subscription, no prompt leakage.

Open in Abliteration