2024.01.08: Initial Test version Release of Solar-Ko
Open-Solar-Ko ⭐🇰🇷
Solar-Ko represents an advanced iteration of the upstage/SOLAR-10.7B-v1.0 model, featuring an expanded vocabulary and the inclusion of a Korean corpus for enhanced pretraining.
As training was conducted solely with publicly available corpora, this model is open for unrestricted use by everyone, adhering to the Apache2.0 open source License.
Model Details
Model Developers: Junbum Lee (Beomi)
Variations: Solar-Ko is available with one parameter sizes — 10B with Continual Pretrained version.
Input: The model accepts only text input.
Output: The model produces text output exclusively.
Model Architecture:
SOLAR-KO-10.7B is an auto-regressive language model that leverages an optimized transformer architecture derived from Llama-2.
Training Data
Parameters
Content Length
GQA
Tokens
Learning Rate
SOLAR-KO-10.7B
A curated mix of Publicly Accessible Korean Corpora
10.7B
2k
✘
>15B*
5e-5
Training Corpus
The model was trained using selected datasets from AIHub and Modu Corpus. Detailed information about the training datasets is available below: