The BFCL-Hi (Hindi BFCL) dataset evaluates the function-calling capability of large language models (LLMs) when questions are asked in Hindi. This is the GCP-translated version of the English BFCL dataset, in which question-function-answer pairs across various domains and multiple languages are originally curated in English.
This dataset is ready for commercial/non-commercial use.
The evaluation steps are described here.
Other Hindi benchmark datasets: [… See the full description on the dataset page:
https://huggingface.co/datasets/nvidia/BFCL-Hi.