CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01minpeter /bfcl-v1-non-live-ast-hermestext1K<n<10K0 likes1.4k downloads2y agoHugging Face02minpeter /bfcl-v1-non-live-ast-parsed [PARSED] BFCL V1 AST (non-live python) The data in this dataset is a subset of the original gorilla-llm/Berkeley-Function-Calling-Leaderboard Subset name multi-turn parallel multiple definition Last turn type number of dataset simple no no no tool_calls 400 multiple no no yes tool_calls 200 parallel no yes no tool_calls 200 parallel_multiple no yes yes tool_calls 200 This is a re-parsing formatting dataset for Python AST parts from V1 of the official dataset of… See the full description on the dataset page: https://huggingface.co/datasets/minpeter/bfcl-v1-non-live-ast-parsed.texttext-generation1K<n<10K1 likes1.1k downloads2y agoHugging Face03teddyyyy123 /bfcl_v3text1K<n<10K0 likes568 downloads2y agoHugging Face04fireworks-ai /bfcl_v3_multi_turn_basetextn<1K1 likes399 downloads2y agoHugging Face05llamastack /bfcl_v3text1K<n<10K7 likes375 downloads1y agoHugging Face06ServiceNow-AI /BFCL_v3_audio License Gorilla is Apache 2.0 licensed, making it suitable for both academic and commercial use. https://github.com/ShishirPatil/gorilla/blob/main/LICENSE Citation @article{patil2023gorilla, title={Gorilla: Large Language Model Connected with Massive APIs}, author={Shishir G. Patil and Tianjun Zhang and Xin Wang and Joseph E. Gonzalez}, year={2023}, journal={arXiv preprint arXiv:2305.15334}, } audio1K<n<10K1 likes323 downloads1y agoHugging Face07hjshah /bfcl_v3text1K<n<10K1 likes288 downloads1y agoHugging Face08muradil211 /ToolWeave_BFCL_Rollout_Case_Study 🧵 ToolWeave BFCL Formal-Training Rollout Case Study This dataset publishes the complete raw on-policy rollout artifact from ToolWeave formal-training update 2, together with a focused real-rollout case study and deterministic K=16 peer-group analysis for runtime-interaction credit assignment. The records contain protocol failures and self-correction; they are raw reinforcement-learning trajectories, not curated demonstrations and not benchmark results. Project:… See the full description on the dataset page: https://huggingface.co/datasets/muradil211/ToolWeave_BFCL_Rollout_Case_Study.tabularn<1K1 likes169 downloads29d agoHugging Face09dvilasuero /bfcl bfcl Evaluation Results Eval created with evaljobs. This dataset contains evaluation results for the model(s) hf-inference-providers/meta-llama/Llama-3.1-8B-Instruct using the eval inspect_evals/bfcl from Inspect Evals. To browse the results interactively, visit this Space. Command This eval was run with: evaljobs inspect_evals/bfcl \ --model hf-inference-providers/meta-llama/Llama-3.1-8B-Instruct \ --name bfcl Run with other models To run this… See the full description on the dataset page: https://huggingface.co/datasets/dvilasuero/bfcl.tabularn<1K0 likes158 downloads10mo agoHugging Face10hjshah /bfcl_v2_asttext1K<n<10K0 likes152 downloads1y agoHugging Face11DCAgent3 /bfcl_parity_a3_rl_DCAgent_selfinstruct_naive_sandboxes_2_verified_70_8B_20260604_192648textn<1K0 likes152 downloads4mo agoHugging Face12hjshah /bfcl_v2_pythontext1K<n<10K0 likes150 downloads1y agoHugging Face13DCAgent2 /DCAgent2_bfcl-parity_laion_GLM-4_7-inferredbugs-sandboxes-maxeps-131k_20260226_044018textn<1K0 likes116 downloads7mo agoHugging Face14DCAgent3 /bfcl_parity_GLM_4_7_swesmith_sandboxes_with_tests_oracle_verified_120s_maxeps_1015c97b3textn<1K0 likes100 downloads4mo agoHugging Face15DCAgent3 /bfcl_parity_a3_rl_DCAgent_inferredbugs_sandboxes_verifier_55_8B_20260526_214919textn<1K0 likes97 downloads4mo agoHugging Face16RioLee /TRBench-BFCL TRBench-BFCL [Paper] | [Model] | [Benchmark] | [Code] 💡 Summary This dataset is a part of ToolRM: Towards Agentic Tool-Use Reward Modeling and serves as a dedicated benchmark for evaluating reward models in tool-use settings. It comprises 2,983 preference annotations buit upon BFCL V3, with assistant responses extracted from archived trajectories available in this github repo. 🌟 Overview ToolRM is a family of lightweight generative and… See the full description on the dataset page: https://huggingface.co/datasets/RioLee/TRBench-BFCL.tabulartext-classification10K<n<100K4 likes82 downloads8mo agoHugging Face17RZ412 /bfcl_multi_turn_datasettextn<1K0 likes77 downloads2y agoHugging Face18clementepasti /bfcl-activations-fulloutputBFCL probe activations, answer-INCLUSIVE (prompt+think+answer prefill), qwen3-8b, dense layers 10-34, last-token + mean pooling. acts_full_output/ = 1,800-problem split, acts_full_output_extra/ = 1,841 complement. kfold_ansinc/ = K=3 answer-inclusive linear classifiers (layer 28) deployed in live best-of-100 (condition C). Rebuild: experiments/bfcl_cot_clf/ in github.com/genlm/rollouts (clement/wip). text100K<n<1M0 likes76 downloads2mo agoHugging Face19minh132 /bfcltext1K<n<10K0 likes71 downloads1y agoHugging Face20hjshah /bfcl_v3_10eachtextn<1K0 likes71 downloads1y agoHugging Face21ezosa /bfcl-nonlivetext1K<n<10K0 likes70 downloads18d agoHugging Face22ezosa /bfcl-livetext1K<n<10K0 likes65 downloads17d agoHugging Face23ezosa /bfcltext1K<n<10K0 likes65 downloads17d agoHugging Face24hjshah /bfcl_v3_only_multi_turntextn<1K0 likes64 downloads1y agoHugging Face25hjshah /bfcltext1K<n<10K0 likes62 downloads1y agoHugging Face26DynaGuard /bfcltext10K<n<100K0 likes59 downloads7mo agoHugging Face27jeypiii /bfcl_v4_single_turntext1K<n<10K0 likes46 downloads6mo agoHugging Face28llamastack /bfcl_v3__oldtext1K<n<10K0 likes43 downloads1y agoHugging Face29hjshah /bfcl_v2_non_pythontextn<1K0 likes41 downloads1y agoHugging Face30hjshah /bfcl_v3_apitextn<1K0 likes38 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.