MicroLlama-134M-Instruct is a custom-trained Small Language Model (SLM) created by Abhiray (Jay). Built using a scaled-down Llama architecture, this model is designed to be highly efficient, lightweight, and can try to conversational instruction-following.
The model strictly follows the ChatML-style template used during its SFT phase. For optimal performance, a generation temperature between 0.3 and 0.5 with a gentle repetition_penalty (e.g., 1.05) is recommended.
1<|system|>
2You are a highly capable, friendly, and helpful AI assistant.</s>
3<|user|>
4What is the core temperature of the Sun?</s>
5<|assistant|>