This is an abliterated version of Meta's Llama-3.1-8B-Instruct model, modified to reduce harmful outputs while maintaining general performance.
This model uses activation-based ablation techniques to modify the model's behavior regarding potentially harmful content. The technique involves:
While maintaining improved safety characteristics compared to the base model.
This model aims to reduce potentially harmful outputs while maintaining functionality. However, users should:
@misc{llama-3.1-8b-instruct-abliterated,
author = {[Your Name]},
title = {Llama-3.1-8B-Instruct-abliterated},
year = {2024},
publisher = {Hugging Face},
journal = {Hugging Face Model Hub},
}