Developed by: Within Us AI
Foundation dataset emphasizing tests-as-truth, agentic loops, and evaluation thinking.
Tests-as-truth supervision patterns
Diff-first patching
Agentic loops (plan→edit→test→reflect) with bounded budgets
Tool-call trace supervision (where present)
Governance/audit & policy-gate awareness
Parquet unavailable (No module named… See the full description on the dataset page:
https://huggingface.co/datasets/WithinUsAI/Genesis_AI_Code_10k.