This is a supervised fine-tuned 125M-parameter causal language model based on
thesreedath/slm-125m-base.
Task types include grounded question answering, unanswerable/refusal cases,
multi-step QA, JSON extraction, summarization, plain-English rewrites, and
comparative QA.
This is a small research/experimentation model for legal and financial
instruction-following behavior. It is not a substitute for professional legal,
financial, or compliance advice.
1<|system|>
2You are a careful legal and financial assistant...
3<|user|>
4Context:
5...
6
7Question:
8...
9<|assistant|>
10...
11<|eos|>