datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
earnings-calls-qa
Lamini Earning Calls QA Dataset
Description
This dataset contains transcripts of earning calls for various companies, along with questions and answers related to the companies' financial performance and other relevant topics.
Format
The transcripts, questions, and answers are in the form of jsonlines files, with each json object in the file containing the transcript of an earning call for a single company.
Data Pipeline Code
The entire data pipeline… See the full description on the dataset page: https://huggingface.co/datasets/lamini/earnings-calls-qa.unified-tool-calls
unified-tool-calls
A single consolidated corpus of tool-calling conversations converted from four source datasets into one unified format.
Source datasets
source
repository
raw rows
converted
in final corpus
xlam
dusersad12/xlam-function-calling-60k
100
97
92
toolace
dusersad12/ToolACE
30
30
28
glaive
dusersad12/glaive_toolcall_en
100
97
92
hermes
dusersad12/hermes-tool-calls
18
18
16
Total entries in the merged corpus: 228.… See the full description on the dataset page: https://huggingface.co/datasets/dusersad12/unified-tool-calls.func_calls
retrain-pipelines Function Calling
version 0.237 - 2026-08-25 12:39:39 UTC
Source datasets :
main :
Xlam Function Calling 60k
Salesforce/xlam-function-calling-60k
(26d14eb -
2025-01-24 19:25:58 UTC)
license :
cc-by-4.0
arxiv :
- 2406.18518
data-enrichment :
Natural Questions Clean
lighteval/natural_questions_clean
(a72f7fa -
2023-10-17 20:29:08 UTC)
license :
unknown
The herein dataset has 2 configs :… See the full description on the dataset page: https://huggingface.co/datasets/retrain-pipelines/func_calls.func_calls_ds
retrain-pipelines Function Calling
version 0.44 - 2026-03-01 12:02:11 UTC
Source datasets :
main :
Xlam Function Calling 60k
Salesforce/xlam-function-calling-60k
(26d14eb -
2025-01-24 19:25:58 UTC)
license :
cc-by-4.0
arxiv :
- 2406.18518
data-enrichment :
Natural Questions Clean
lighteval/natural_questions_clean
(a72f7fa -
2023-10-17 20:29:08 UTC)
license :
unknown
The herein dataset has 2 configs : continued_pre_training and supervised_finetuning.
The former… See the full description on the dataset page: https://huggingface.co/datasets/retrain-pipelines/func_calls_ds.tool-calls-mini
tool-calls-mini
500 synthetic tool-calling conversations in TRL's conversational format,
for supervised fine-tuning. Built to be coherent: every tool result is a plausible
function of the arguments it was called with, and every final answer reflects that
result — so the set teaches when to call a tool, not just what a call looks like.
Format
Each row has messages and tools. An assistant turn carries tool_calls instead of
content; the tool replies as a tool role… See the full description on the dataset page: https://huggingface.co/datasets/qgallouedec/tool-calls-mini.urdu-emergency-calls
Urdu Emergency Call Conversations Dataset (Pakistan)
Overview
This dataset contains 5,000 curated Urdu emergency call conversation samples from the Pakistan region, designed to support training and evaluation of Urdu Large Language Models (LLMs) for emergency response, command centers, and interpreter-style systems.
The conversations simulate real-world emergency scenarios such as:
Floods
Medical emergencies
Accidents
Crimes
Natural disasters
Public safety… See the full description on the dataset page: https://huggingface.co/datasets/abeeranajam31/urdu-emergency-calls.moh_8_rollouts_tool_calls
MOH 8 Rollouts Tool Calls
Fixed 8-rollout benchmark sampled from aimosprite/training-output using only tool-calling attempts.
Construction
Source problem family: polymath_*
Source eligibility band: correct_count_16 in [1, 10]
Candidate attempt pool per problem: attempts with Python Calls > 0
Final sample per problem: 8 attempts, sampled without replacement using seed 42
Additional constraint: the sampled 8 always include at least one correct attempt
Published rows: 300… See the full description on the dataset page: https://huggingface.co/datasets/aimosprite/moh_8_rollouts_tool_calls.urdu-emergency-calls
Urdu Emergency Call Conversations Dataset (Pakistan)
Overview
This dataset contains 5,000 curated Urdu emergency call conversation samples from the Pakistan region, designed to support training and evaluation of Urdu Large Language Models (LLMs) for emergency response, command centers, and interpreter-style systems.
The conversations simulate real-world emergency scenarios such as:
Floods
Medical emergencies
Accidents
Crimes
Natural disasters
Public safety incidents
The… See the full description on the dataset page: https://huggingface.co/datasets/hamza-amin/urdu-emergency-calls.type-schema-tools-calls
type-schema-tools-calls
Dataset for TypeSchema Tool Calling
