This model is an initial test conducted in the summer of 2025 shortly after the release of Liquid.AI's
LFM2 model.
The idea was to quickly get an opinion on this hybrid convolutional SSM (StripedHyena) & attention architecture (we really love
SSMs and especially
convolutional SSMs, which we believe don't get enough attention compared to recurrent SSMs).
In particular, the resources required to train the compact base model (the 1.2B parameter model can be finetuned on one 80GB A100 without quantization, according to our observations).
As part of this initial test, we fine-tuned the model using the
notebook provided by Liquid.AI on 195,583 non-synthetic French DPO data.