Raw evaluation artifacts backing the intelligencemarkets
project (an interactive visualization of AI model economics across benchmarks).
This dataset holds the heavy raw inputs for the SWE-bench pipeline. The small
inputs (terminal_data.jsonl) and the generated .npz / data.json outputs live
directly in the GitHub repo.
swebench/
evals/
minicoder4b/ # one JSON per eval job (SWE-bench harness output)… See the full description on the dataset page:
https://huggingface.co/datasets/ricdomolm/intelligence-markets-data.