CAPEval (Coverage And Precision Evaluation) is a checklist-based caption evaluation benchmark.
It decouples caption quality into Coverage (C) and Precision (P) (0–100), and studies how each profile transfers to VLM understanding and T2I generation.
Code / docs: liuzhipenggg/CAPEval
Project page: liuzhipenggg.github.io/CAPEval
Paper: arXiv:2608.02589
Leaderboard: leaderboard
image/
300 high-resolution images… See the full description on the dataset page:
https://huggingface.co/datasets/LiuzhipengUCAS/CAPEval.