This repository contains tasks from the Fermatix SWE Bench dataset, intended for evaluating the capabilities of models in automatic bug fixing and code modification.
The tasks cover various programming languages and projects, providing a diverse set of scenarios for testing and training.
Each task includes:
Patches with fixes and/or tests
Instructions for building and running (in the form of a Dockerfile), as well as corresponding run logs
A parquet file… See the full description on the dataset page:
https://huggingface.co/datasets/fermatix-ai/Fermatix-SWE-Bench.