Views
No views yet
llama.cpp, specifically utilizing the imatrix quantization option for potentially improved performance.
system role based on its fine-tuning:
1<start_of_turn>system
2{optional system prompt here}<end_of_turn>
3<start_of_turn>user
4{User messages. You can also place the system prompt here.}<end_of_turn>
5<start_of_turn>model
6{Model's response}<end_of_turn>