Fixberry is a little dataset I have made to train models to correctly count the number of letters in a word. It is commonly known that even the best LLMs fail at counting the number of R's in strawberry. I have also found out they have problems with other words too, like keeper and parallel but weirdly not with words like pepper and peeper. This should really be investigated more closely, I suspect it has something to do with tokenization and the possability that the model… See the full description on the dataset page:
https://huggingface.co/datasets/Khawn2u/Fixberry.