Quantization made by Richard Erkhov.
llama3-8B-DarkIdol-2.0-Uncensored - AWQ
Original model description:
license: llama3
language:
en
ja
zh
tags:
roleplay
llama3
sillytavern
idol
Special Thanks:
Ransss's superb gguf version, thank you for your conscientious and responsible dedication.
https://huggingface.co/Ransss/llama3-8B-DarkIdol-2.0-Uncensored-Q8_0-GGUF
Model Description:
The module combination has been readjusted to better fulfill various roles and has been adapted for mobile phones.
Saving money(LLama 3)
Uncensored
Quick response
A scholarly response akin to a thesis.(I tend to write songs extensively, to the point where one song almost becomes as detailed as a thesis. :)
DarkIdol:Roles that you can imagine and those that you cannot imagine.
Roleplay
Specialized in various role-playing scenarios
more look at test role. (https://huggingface.co/aifeifei798/llama3-8B-DarkIdol-1.2/resolve/main/test )
more look at LM Studio presets (https://huggingface.co/aifeifei798/llama3-8B-DarkIdol-1.2/resolve/main/config-presets )
image/png
Chang Log
2024-06-26
之前版本的迭代太多了,已经开始出现过拟合现象.重新使用了新的工艺重新制作模型,虽然制作复杂了,结果很好,新的迭代工艺如图
The previous version had undergone excessive iterations, resulting in overfitting. We have recreated the model using a new process, which, although more complex to produce, has yielded excellent results. The new iterative process is depicted in the figure.
image/png
16K,32K或更多的context是我下一个版本要做的
The next version I am working on will feature 16K, 32K, or even larger context sizes.
Questions
The model's response results are for reference only, please do not fully trust them.
I am unable to test Japanese and Korean parts very well. Based on my testing, Korean performs excellently, but sometimes Japanese may have furigana (if anyone knows a good Japanese language module, - I need to replace the module for integration).
With the new manufacturing process, overfitting and crashes have been reduced, but there may be new issues, so please leave a message if you encounter any.
testing with other tools is not comprehensive.but there may be new issues, so please leave a message if you encounter any.
问题
模型回复结果仅供参考,请勿完全相信
日语,韩语部分我没办法进行很好的测试,根据我测试情况,韩语表现的很好,日语有时候会出现注音(谁知道好的日文语言模块,我需要换模块集成)
新工艺制作,过拟合现象和崩溃减少了,可能会有新的问题,碰到了请给我留言
其他工具的测试不完善
Stop Strings
1 stop = [
2 "## Instruction:" ,
3 "### Instruction:" ,
4 "<|end_of_text|>" ,
5 " //:" ,
6 "</s>" ,
7 "<3```" ,
8 "### Note:" ,
9 "### Input:" ,
10 "### Response:" ,
11 "### Emoticons:"
12 ] ,
Model Use
Koboldcpp https://github.com/LostRuins/koboldcpp
Since KoboldCpp is taking a while to update with the latest llama.cpp commits, I'll recommend this fork if anyone has issues.
LM Studio https://lmstudio.ai/
llama.cpp https://github.com/ggerganov/llama.cpp
Backyard AI https://backyard.ai/
Meet Layla,Layla is an AI chatbot that runs offline on your device.No internet connection required.No censorship.Complete privacy.Layla Lite https://www.layla-network.ai/
character
If you want to use vision functionality:
You must use the latest versions of Koboldcpp .
To use the multimodal capabilities of this model and use vision you need to load the specified mmproj file, this can be found inside this model repo. Llava MMProj
You can load the mmproj by using the corresponding section in the interface:
image/png
Thank you:
To the authors for their hard work, which has given me more options to easily create what I want. Thank you for your efforts.
Hastagaras
Gryphe
cgato
ChaoticNeutrals
mergekit
merge
transformers
llama
Nitral-AI
MLP-KTLim
rinna
hfl
Rupesh2
stephenlzc
theprint
Sao10K
turboderp
TheBossLevel123
.........
base_model:
Nitral-AI/Hathor_Fractionate-L3-8B-v.05
Hastagaras/Jamet-8B-L3-MK.V-Blackroot
turboderp/llama3-turbcat-instruct-8b
aifeifei798/Meta-Llama-3-8B-Instruct
Sao10K/L3-8B-Stheno-v3.3-32K
TheBossLevel123/Llama3-Toxic-8B-Float16
cgato/L3-TheSpice-8b-v0.8.3
library_name: transformers
tags:
mergekit
merge
llama3-8B-DarkIdol-1.3.1
This is a merge of pre-trained language models created using
mergekit .
Merge Details
Merge Method
This model was merged using the
Model Stock merge method using
aifeifei798/Meta-Llama-3-8B-Instruct as a base.
Models Merged
The following models were included in the merge:
Configuration
The following YAML configuration was used to produce this model:
1 models :
2 - model : Sao10K/L3 - 8B - Stheno - v3.3 - 32K
3 - model : Hastagaras/Jamet - 8B - L3 - MK.V - Blackroot
4 - model : cgato/L3 - TheSpice - 8b - v0.8.3
5 - model : Nitral - AI/Hathor_Fractionate - L3 - 8B - v.05
6 - model : TheBossLevel123/Llama3 - Toxic - 8B - Float16
7 - model : turboderp/llama3 - turbcat - instruct - 8b
8 - model : aifeifei798/Meta - Llama - 3 - 8B - Instruct
9 merge_method : model_stock
10 base_model : aifeifei798/Meta - Llama - 3 - 8B - Instruct
11 dtype : bfloat16
12
base_model:
hfl/llama-3-chinese-8b-instruct-v3
rinna/llama-3-youko-8b
MLP-KTLim/llama-3-Korean-Bllossom-8B
library_name: transformers
tags:
mergekit
merge
llama3-8B-DarkIdol-1.3.2
This is a merge of pre-trained language models created using
mergekit .
Merge Details
Merge Method
This model was merged using the
Model Stock merge method using ./llama3-8B-DarkIdol-1.3.1 as a base.
Models Merged
The following models were included in the merge:
Configuration
The following YAML configuration was used to produce this model:
1 models :
2 - model : hfl/llama - 3 - chinese - 8b - instruct - v3
3 - model : rinna/llama - 3 - youko - 8b
4 - model : MLP - KTLim/llama - 3 - Korean - Bllossom - 8B
5 - model : ./llama3 - 8B - DarkIdol - 1.3.1
6 merge_method : model_stock
7 base_model : ./llama3 - 8B - DarkIdol - 1.3.1
8 dtype : bfloat16
9
base_model:
theprint/Llama-3-8B-Lexi-Smaug-Uncensored
Rupesh2/OrpoLlama-3-8B-instruct-uncensored
stephenlzc/dolphin-llama3-zh-cn-uncensored
library_name: transformers
tags:
mergekit
merge
llama3-8B-DarkIdol-2.0-Uncensored
This is a merge of pre-trained language models created using
mergekit .
Merge Details
Merge Method
This model was merged using the
Model Stock merge method using ./llama3-8B-DarkIdol-1.3.2 as a base.
Models Merged
The following models were included in the merge:
Configuration
The following YAML configuration was used to produce this model:
1 models :
2 - model : Rupesh2/OrpoLlama - 3 - 8B - instruct - uncensored
3 - model : stephenlzc/dolphin - llama3 - zh - cn - uncensored
4 - model : theprint/Llama - 3 - 8B - Lexi - Smaug - Uncensored
5 - model : ./llama3 - 8B - DarkIdol - 1.3.2
6 merge_method : model_stock
7 base_model : ./llama3 - 8B - DarkIdol - 2.0 - Uncensored
8 dtype : bfloat16
9