datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
browse_compaguvis-stage-2aguvis-stage-1post-train-bench-traces
PostTrainBench Sessions by Benchmark
Derived from akseljoonas/posttrainbench-sessions on 2026-04-20.
This dataset exports each source row as one viewer-compatible JSONL trace and groups traces by benchmark.
Layout
benchmarks.json: benchmark catalog and counts
benchmarks/<benchmark>/index.json: metadata index for one benchmark
benchmarks/<benchmark>/<job_id>.jsonl: one converted session trace per source row
Benchmarks
Benchmark
Sessions
aime2025
19… See the full description on the dataset page: https://huggingface.co/datasets/smolagents/post-train-bench-traces.GAIA-annotatedandroid-controltoolcallinggaia-tracesguiact-web-singleaguvis-stage-testhermes-codeagenthermes-function-calling-v1-formatted-code-agentcodeagent-tracessmolagentssmolagents-toolcalling-mergedtraining-tracessmolagents-mergedresultssynthetic-tracesglaive-function-calling-with-reasoningsmolagents-merged-filteredbenchmark-v1SecretAgenda_Game_Smolagents_Gemmascope_SAE_Datatool-scrapingsmolagents_benchmark_200smoltalk2_smolagents_toolcalling_french
Description
This is the SFT/smolagents_toolcalling_traces_think subset of HuggingFaceTB/smoltalk2, a tool calling dataset.We've translated the prompt and final anwser into French, while the rest (the tool call trace) remains in English.
CodeARC-Problems-TestCases-Filteredtrace-generation-taskssmol_agents_benchmark_300answers
