MeowBench is a high-fidelity, expert-verified quad-modal benchmark designed to evaluate Multimodal Large Language Models (MLLMs) on feline intention decoding. It is the official evaluation suite for the Meow-Omni 1 model.
MeowBench is designed to solve the challenge of "semantic aliasing" in animal behaviour. It provides a rigorous testing ground for models to determine if they can move beyond superficial pattern matching to… See the full description on the dataset page:
https://huggingface.co/datasets/smgjch/MeowBench.