Layer 18 of the original model was chosen for abliteration.
I also created another layer 17 abliterated model for comparison.
These two layers were chosen due to they both produce uncensored response
after respective layer was abliterated.
It is uploaded here to be evaluated by the Open LLM Leaderboard to see how brain damaged it
is compared to the original model.
ORPO fine tuning is currently underway to see if it can regain its sanity. You can play with this model first or wait until I am done with the fine tuning.
Benchmark (100.0*raw scores only)
Click on the model name go to the raw score json generated by Open LLM Leaderboard.