Blind Spots of Qwen3-0.6B-Base
Description
This dataset contains 10 examples of inputs where the Qwen3-0.6B-Base model produced incorrect or unexpected outputs. The goal is to document "blind spots" of the model across different reasoning, logic, and knowledge tasks. Each row includes the input given to the model, the expected_output, the model_output, and the error_type.
Columns in the dataset:
input: The text prompt given to the model.
expected_output: The correct or intended response.… See the full description on the dataset page:
https://huggingface.co/datasets/Abdiaziz01/frontier_model_blind_spots.