nvidia/BFCL-Hi
Dataset Description: The BFCL-Hi (Hindi BFCL) dataset evaluates the function-calling capability of large language models (LLMs) when questions are asked in Hindi. This is the GCP-translated version of the English BFCL dataset, in which question-function-answer pairs across various domains and multiple languages are originally curated in English. This dataset is ready for commercial/non-commercial use. The evaluation steps are described here. Other Hindi benchmark datasets: [… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/BFCL-Hi.
3378
1---2license: cc-by-4.03task_categories:4- text-generation5language:6- hi7tags:8- BFCL9- Hindi10pretty_name: Hindi BFCL11size_categories:12- 1K<n<10K13---14 15## Dataset Description:16The BFCL-Hi (Hindi BFCL) dataset evaluates the function-calling capability of large language models (LLMs) when questions are asked in Hindi. This is the GCP-translated version of the English BFCL dataset, in which question-function-answer pairs across various domains and multiple languages are originally curated in English.17 18This dataset is ready for commercial/non-commercial use.19The evaluation steps are described [here](https://huggingface.co/datasets/nvidia/BFCL-Hi/blob/main/EVAL.md). <br>20Other Hindi benchmark datasets: [ [IFEval-Hi](https://huggingface.co/datasets/nvidia/IFEval-Hi) | [MT-Bench-Hi](https://huggingface.co/datasets/nvidia/MT-Bench-Hi) | [GSM8K-Hi](https://huggingface.co/datasets/nvidia/GSM8K-Hi) | [BFCL-Hi](https://huggingface.co/datasets/nvidia/BFCL-Hi) | [ChatRAG-Hi](https://huggingface.co/datasets/nvidia/ChatRAG-Hi) ]21 22## Dataset Owner:23NVIDIA Corporation24 25 26## Dataset Creation Date:27April 202528 29## License/Terms of Use: 30This dataset is licensed under the Creative Commons Attribution 4.0 International License (CC-BY-4.0). Additional Information: Apache312.0 License.32 33## Intended Usage:34Evaluate LLM's ability to call functions and tools when queries are asked in the Hindi language.35 36## Dataset Characterization37Data Collection Method<br>38* Synthetic<br>39 40Labeling Method<br>41* Not Applicable <br>42 43## Dataset Format44Text45 46## Dataset Quantification476.9MB of prompt-label pairs, comprising 2251 individual samples.48 49## Ethical Considerations:50NVIDIA believes Trustworthy AI is a shared responsibility, and we have established policies and practices to enable development for a wide array of AI applications. When downloaded or used in accordance with our terms of service, developers should work with their internal model team to ensure this model meets requirements for the relevant industry and use case and addresses unforeseen product misuse. 51 52Please report security vulnerabilities or NVIDIA AI Concerns [here](https://www.nvidia.com/en-us/support/submit-security-vulnerability/).53 54## Citing55 56If you find our work helpful, please consider citing our paper:57```58@article{kamath2025benchmarking,59 title={Benchmarking Hindi LLMs: A New Suite of Datasets and a Comparative Analysis},60 author={Kamath, Anusha and Singla, Kanishk and Paul, Rakesh and Joshi, Raviraj and Vaidya, Utkarsh and Chauhan, Sanjay Singh and Wartikar, Niranjan},61 journal={arXiv preprint arXiv:2508.19831},62 year={2025}63}64```65 