The 26-op corpus, experiment drivers, and analysis scripts behind the four
gpuemu preprints, led by "The Correctness Illusion in LLM-Generated GPU
Kernels" (arXiv:2606.20128). A
controlled set for measuring whether a correctness oracle actually catches the
bugs LLM-generated GPU kernels routinely contain — plus the full harness that
produces every table and figure in the papers.