HumanEval Benchmark for Gleam and MoonBit
Dataset Description
This dataset contains a translation of the HumanEval benchmark into two no-resource programming languages: Gleam and MoonBit.
This dataset is derived from the MultiPL-E benchmark, which extends HumanEval by translating its programming tasks into multiple programming languages.
More details on this benchmark and how it was created can be found in the paper No Resource, No Benchmarks, No Problem?… See the full description on the dataset page: https://huggingface.co/datasets/Devy1/humaneval-no-resource.