๐ We present FERMAT, a benchmark designed to rigorously test the multimodal reasoning and auto-evaluation capabilities of Vision-Language Models (VLMs) using real-world handwritten math problems! ๐๐ง
2,244 handwritten math solutions, carefully annotated across core mathematical domains:
๐ข Arithmetic | ๐ Algebra | ๐ Geometry | ๐ Mensuration | ๐ฒ Probability | ๐ Statistics | ๐ Trigonometry | ๐โฆ See the full description on the dataset page:
https://huggingface.co/datasets/ai4bharat/FERMAT.