Version: 0.1 It can't be your daily driver yet
The "true" tiny
Gemma 3 270M can't really chat. So this is an attempt at getting a decent (multi turn) chat model out of it. Also an experiment in dataset curation, building fine-tuning pipelines and seeing where a 270M would get saturated/can't improve further without regressing on important bits.
This is the MLX variant for running on Apple silicon (Mac's, iPhone's, and iPads). For other platforms you can
use the LoRA to merge and create. I will upload a GGUF verions later.
Local Chat model for personal use.
Before:
'''
Okay, I'm ready to explain different parts of reasoning! To give you the best explanation, please tell me what type of reasoning you'd like to understand. For example:
-
Logical deduction: This involves making deductions from general premises to arrive at a specific conclusion. It involves reasoning from general principles to specific instances.
-
Deductive reasoning: This involves making logical deductions from specific premises to arrive at a general conclusion. It involves reasoning from general principles to specific instances.
-
Inductive reasoning: This involves making general conclusions from specific instances. It involves reasoning from specific instances to a general conclusion.
-
Inductive reasoning: This involves making general conclusions from specific instances.
'''