Joint text+image embeddings for every card in Slay the Spire 2 (Early Access), produced by Qwen/Qwen3-VL-Embedding-2B. One unit-normalized 1024-D vector per card. Mechanically AND visually similar cards land near each other; cards across STS1 and STS2 share the coordinate system.
This is the multimodal-embeddings dataset. For text-only embeddings or the underlying card metadata + portraits, see:
t22000t/slay-the-spire-2-cards - metadata… See the full description on the dataset page:
https://huggingface.co/datasets/t22000t/slay-the-spire-2-card-multimodal-embeddings.