Beta
Explore
Marketplace
Neural Labs
Chat
Wallet
Docs
Unholy-8B-DPO-OAS – AI Model by Undi95 | AlphaNeural AI
You can deploy this model and start earning money today!
Undi95
/
Unholy-8B-DPO-OAS
like
0
transformers
safetensors
llama
text-generation
conversational
autotrain_compatible
text-generation-inference
endpoints_compatible
us
Views
No views yet
Model card
Files and Versions
Community
API
Deploy
This is a TEST It was made with a custom Orthogonal Activation Steering script I shared HERE :
https://huggingface.co/posts/Undi95/318385306588047#663609dc1818d469455c0222
(but be ready to put your hands in some fucked up code bro)
Step :
First I took Unholy (FT of L3 on Toxic Dataset)
Then I trained 2 epoch of DPO on top, with the SAME dataset (
https://wandb.ai/undis95/Uncensored8BDPO/runs/3rg4rz13/workspace?nw=nwuserundis95
)
Finally, I used OAS on top, bruteforcing the layer to get the best one (I don't really understand all of this, sorry)