"i1" prefix: Indicates imatrix was used during quantization
Original Model
Gumini-1.5B (구미니) is a bilingual Korean-English base language model trained using the Inheritune methodology. Starting from Qwen 2.5 3B, the model progressively grew from 10 to 16 layers through 7 training stages.
This model is Built with Qwen and derived from Qwen 2.5 3B.
Qwen is licensed under the Qwen RESEARCH LICENSE AGREEMENT.
Copyright (c) Alibaba Cloud. All Rights Reserved.
This model is for NON-COMMERCIAL / RESEARCH use only.
For commercial use, contact Alibaba Cloud.
References
Inheritune Paper
bibtex
1@inproceedings{Sanyal2024inheritune,
2 title={Inheritune: Training Smaller Yet More Attentive Language Models},
3 author={Sunny Sanyal and Ravid Shwartz-Ziv and Alexandros G. Dimakis and Sujay Sanghavi},
4 year={2024},
5 url={https://arxiv.org/abs/2404.08634}
6}
Qwen 2.5
bibtex
1@misc{qwen2.5,
2 title={Qwen2.5: A Party of Foundation Models},
3 author={Qwen Team},
4 year={2024},
5 url={https://qwenlm.github.io/blog/qwen2.5/}
6}
Citation
bibtex
1@misc{gumini2025,
2 title={Gumini-1.5B: Bilingual Korean-English Language Model via Inheritune},
3 author={Gumin Kwon},
4 year={2025},
5 note={Built with Qwen. Trained with Inheritune progressive layer growing.},
6 url={https://huggingface.co/GuminiResearch/Gumini-1.5B-Base-i1-GGUF}
7}