This repository contains the train split of MVEB (Multimodal Visual identity Embedding Benchmark) — a benchmark for identity-level retrieval. Given a query (text + image), a model retrieves candidates (text + image) that belong to the same identity.
The train split covers 20 subsets across four meta-tasks:
Identity Recognition — object / product / species recognition
Re-Identification — person / face / vehicle re-ID
Identity Grounding —… See the full description on the dataset page:
https://huggingface.co/datasets/HugC/MVEB-train.