A dataset for frame-aware intelligence.
Current LLMs often answer inside a broken question, reinforcing:
This benchmark evaluates the ability to stop, and clarify the premise before responding.
clarify
Identify instability in the prompt and restate what must be resolved first
There is no “answering” action in this… See the full description on the dataset page:
https://huggingface.co/datasets/ClarusC64/ecb_v01.