datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dbbench_cot_enriched_for_agentbenchagentbench-sft-trajectories-v4-longexp-fullDARA-Agentbench
Dataset Card for DARA-Agentbench
Dataset Summary
This dataset contains 577 curated reasoning trajectories for KGQA LLM-based agents in the Agentbench format. (https://github.com/UKPLab/acl2024-DARA). It is sourced from GrailQA, WebQSP, and GraphQ.
The fields include:
raw question: original question
input: question with the linked entities.
output: The step-by-step reasoning trajectory to construct the full logical form.
variable list: The stepwise logical forms… See the full description on the dataset page: https://huggingface.co/datasets/UKPLab/DARA-Agentbench.agentbench-sft-trajectories-v2-dbexpandagentbench-sft-trajectories-v3-planA
agentbench-sft-trajectories-v3-planA
Merged dataset for LLM Advanced Competition SFT (Plan A).
Composition
Source
Records
Pct
u-10bei/dbbench_sft_dataset_react (v1)
300
5.5%
u-10bei/dbbench_sft_dataset_react_v2
360
6.6%
u-10bei/dbbench_sft_dataset_react_v3
1,112
20.3%
u-10bei/dbbench_sft_dataset_react_v4
1,194
21.8%
u-10bei/sft_alfworld_trajectory_dataset_v5
2,502
45.8%
Total (after dedup)
5,468
100%
Design
All 4 versions of… See the full description on the dataset page: https://huggingface.co/datasets/sabia0080/agentbench-sft-trajectories-v3-planA.agentbench-sft-trajectories-v2-enhancedagentbench-sft-trajectories-v2agentbench-sft-trajectories-v1
