datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
autotrainer-v0
autotrainer-v0
AutoTrainer-v0: LLM-agent-controlled GRPO training on Countdown
Dataset Info
Rows: 1
Columns: 1
Columns
Column
Type
Description
state_json
Value('string')
Full autotrainer state as JSON string
Generation Parameters
{
"script_name": "run_round.py",
"model": "Qwen/Qwen2.5-1.5B-Instruct",
"description": "AutoTrainer-v0: LLM-agent-controlled GRPO training on Countdown",
"experiment_id": "autotrainer-v0"… See the full description on the dataset page: https://huggingface.co/datasets/raca-workspace-v1/autotrainer-v0.autotrainer-v1
AutoTrainer v1
Agentic training harness where Claude decides every training step's method, data, and hyperparameters.
Configs
steps: Per-step decisions, metrics, agent traces, and costs
eval_traces: Per-question evaluation results with difficulty breakdown (n_args=2-10)
autotrainer-v1-run-v2-11stepsautotrain-data-8a00-9wrj-ig4m
Dataset Card for "autotrain-data-8a00-9wrj-ig4m"
More Information needed
autotrain-data-Nuclear_Fusion_Falcon
Dataset Card for "autotrain-data-Nuclear_Fusion_Falcon"
More Information needed
ordis-autotrain-pure-skill
Ordis Pure Skill Features for AutoTrain
Dataset for boat racing prediction using only skill-based features (no odds).
Files
train.parquet: Training data (21,684 samples)
valid.parquet: Validation data (28,316 samples)
Features
221 pure skill features
Binary target: label (1=top-3 finish, 0=other)
Usage with AutoTrain
Go to https://huggingface.co/autotrain
Create new project → Tabular → Binary Classification
Use dataset:… See the full description on the dataset page: https://huggingface.co/datasets/sugiken/ordis-autotrain-pure-skill.autotrainer-v0-eval-tracesautotrainer-v1-v2abhishek__autotrain-llama3-70b-orpo-v1details_abhishek__autotrain-llama3-70b-orpo-v1
Dataset Card for Evaluation run of abhishek/autotrain-llama3-70b-orpo-v1
Dataset automatically created during the evaluation run of model abhishek/autotrain-llama3-70b-orpo-v1.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_abhishek__autotrain-llama3-70b-orpo-v1.abhishek__autotrain-llama3-orpo-v2details_abhishek__autotrain-llama3-70b-orpo-v2
Dataset Card for Evaluation run of abhishek/autotrain-llama3-70b-orpo-v2
Dataset automatically created during the evaluation run of model abhishek/autotrain-llama3-70b-orpo-v2.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_abhishek__autotrain-llama3-70b-orpo-v2.
