Dataset for probing model preferences for linguistically acceptable sentences.
Generated by introducing automatic corruptions into sentences from Wikipedia, based on UniMorph minimal tag pairs.
More info coming soon!
@misc{glocker2025growmergescalingstrategies,
title={Grow Up and Merge: Scaling Strategies for Efficient Language Adaptation},
author={Kevin Glocker and Kätriin Kukk and Romina Oji and Marcel Bollmann and Marco Kuhlmann and Jenny Kunz},
year={2025}… See the full description on the dataset page:
https://huggingface.co/datasets/liu-nlp/blimp-single-error.