This dataset is a standardized version of the original benchmark, prepared for easy evaluation of LLMs on planning tasks.
{
"id": "Value(dtype='string', id=None)",
"query": "Value(dtype='string', id=None)",
"org": "Value(dtype='string', id=None)",
"dest":… See the full description on the dataset page:
https://huggingface.co/datasets/tuandunghcmut/travelplanner-eval.