This llama model was trained 2x faster with
Unsloth and Huggingface's TRL library.
環境:Google Colab notebooks
試行コードファイル:LoRA_template_unsloth.20241127.ipynb
ベースモデル:llm-jp/llm-jp-3-13b
学習データセット:ichikara-instruction-003-001.1.json
学習後新モデル名:llm-jp-3-13b-it
模試問題データセット:elyza-tasks-100-TV_0.jsonl
出力:llm-jp-3-13b-it_output.jsonl
スコア:3.01