This dataset was created as part of a technical challenge to identify and document the "blind spots" of a recent, moderately-sized base model. The goal was to browse models released in the last 6 months (between 0.6B and 6B parameters), select one, and systematically probe its failures to understand its limitations.
This dataset contains 10 examples where the Nanbeige/Nanbeige4.1-3B… See the full description on the dataset page:
https://huggingface.co/datasets/abdulrahman245/nanbeige-ai-errors.