A large-scale dataset curated for training and evaluating code generation models. This dataset contains high-quality code snippets, prompts, and metadata suitable for various code synthesis tasks, including prompt completion, function generation, and docstring-to-code translation.
✅ Prompts describing coding tasks
✅ Code solutions in Python (or other languages, if applicable)
✅ Metadata… See the full description on the dataset page:
https://huggingface.co/datasets/Vynh/code-generation-dataset.