datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
math-sft-mixedKronos-Dataset
Kronos-Dataset
It is a collection of various datasets to expand the capabilities of reasoning models in agent tasks, medical reasoning, multilingual thinking, and writing. All of these are unified in a single format:
[
{
"from": "system",
"value": "You are a medical AI assistant with advanced reasoning capabilities. Provide detailed, step-by-step analysis for medical questions."
},
{
"from": "human",
"value": "Given the symptoms of sudden weakness in the left arm and leg, recent… See the full description on the dataset page: https://huggingface.co/datasets/FredyRivera-dev/Kronos-Dataset.T145__KRONOS-8B-V6-details
Dataset Card for Evaluation run of T145/KRONOS-8B-V6
Dataset automatically created during the evaluation run of model T145/KRONOS-8B-V6
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/T145__KRONOS-8B-V6-details.adaption-football-tactics-qa
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
adaption-football_tactics_qa
This instruction-tuning dataset comprises prompt and completion pairs focused on association football and sports analytics. Content covers match analysis, tactical formations, player performance metrics, official rules, and strategic gameplay scenarios. The dataset supports conversational AI training by providing clear, well-reasoned responses ranging from general… See the full description on the dataset page: https://huggingface.co/datasets/kronos17/adaption-football-tactics-qa.T145__KRONOS-8B-V7algebra-sft-1mkronos-r3b-dpo-pairs
Kronos R3 Branch B — DPO preference pairs
11 {prompt, chosen, rejected, question_id} pairs sampled from Qwen3-Coder-480B-A35B-Instruct (temperature=0.7, 4 candidates per problem) on LCB-medium-stdin problems, graded via _grade_public against public test cases.
chosen: shortest-passing teacher candidate
rejected: first-failing teacher candidate
filter: keep only problems where ≥1 candidate passes AND ≥1 fails
Yield notes (cycle-114/115)
30 problems sampled → 11 valid… See the full description on the dataset page: https://huggingface.co/datasets/jaivial/kronos-r3b-dpo-pairs.finMME_procT145__KRONOS-8B-V3T145__KRONOS-8B-V3-details
Dataset Card for Evaluation run of T145/KRONOS-8B-V3
Dataset automatically created during the evaluation run of model T145/KRONOS-8B-V3
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/T145__KRONOS-8B-V3-details.T145__KRONOS-8B-V4-details
Dataset Card for Evaluation run of T145/KRONOS-8B-V4
Dataset automatically created during the evaluation run of model T145/KRONOS-8B-V4
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/T145__KRONOS-8B-V4-details.T145__KRONOS-8B-V8-details
Dataset Card for Evaluation run of T145/KRONOS-8B-V8
Dataset automatically created during the evaluation run of model T145/KRONOS-8B-V8
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/T145__KRONOS-8B-V8-details.T145__KRONOS-8B-V9-details
Dataset Card for Evaluation run of T145/KRONOS-8B-V9
Dataset automatically created during the evaluation run of model T145/KRONOS-8B-V9
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/T145__KRONOS-8B-V9-details.T145__KRONOS-8B-V1-P2-details
Dataset Card for Evaluation run of T145/KRONOS-8B-V1-P2
Dataset automatically created during the evaluation run of model T145/KRONOS-8B-V1-P2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/T145__KRONOS-8B-V1-P2-details.T145__KRONOS-8B-V2-details
Dataset Card for Evaluation run of T145/KRONOS-8B-V2
Dataset automatically created during the evaluation run of model T145/KRONOS-8B-V2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/T145__KRONOS-8B-V2-details.T145__KRONOS-8B-V5-details
Dataset Card for Evaluation run of T145/KRONOS-8B-V5
Dataset automatically created during the evaluation run of model T145/KRONOS-8B-V5
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/T145__KRONOS-8B-V5-details.T145__KRONOS-8B-V7-details
Dataset Card for Evaluation run of T145/KRONOS-8B-V7
Dataset automatically created during the evaluation run of model T145/KRONOS-8B-V7
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/T145__KRONOS-8B-V7-details.T145__KRONOS-8B-V2T145__KRONOS-8B-V6T145__KRONOS-8B-V1-P3-details
Dataset Card for Evaluation run of T145/KRONOS-8B-V1-P3
Dataset automatically created during the evaluation run of model T145/KRONOS-8B-V1-P3
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/T145__KRONOS-8B-V1-P3-details.kronos-distill-seedTemporal_ConversationThis dataset contains various dialogues which contain some temporal sense within them.
It is derived from the TIMEDIAL dataset.
T145__KRONOS-8B-V4T145__KRONOS-8B-V5kronos-r3b-dpo-pairs-hardT145__KRONOS-8B-V1-P1-details
Dataset Card for Evaluation run of T145/KRONOS-8B-V1-P1
Dataset automatically created during the evaluation run of model T145/KRONOS-8B-V1-P1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/T145__KRONOS-8B-V1-P1-details.finMME_embed
