datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
verdclm-llama3-tokenized-shuffledx_dataset_07096
Bittensor Subnet 13 X (Twitter) Dataset
Miner Data Compliance Agreement
In uploading this dataset, I am agreeing to the Macrocosmos Miner Data Compliance Policy.
Dataset Summary
This dataset is part of the Bittensor Subnet 13 decentralized network, containing preprocessed data from X (formerly Twitter). The data is continuously updated by network miners, providing a real-time stream of tweets for various analytical and machine learning… See the full description on the dataset page: https://huggingface.co/datasets/zephyr-1111/x_dataset_07096.x_dataset_0708150
Bittensor Subnet 13 X (Twitter) Dataset
Miner Data Compliance Agreement
In uploading this dataset, I am agreeing to the Macrocosmos Miner Data Compliance Policy.
Dataset Summary
This dataset is part of the Bittensor Subnet 13 decentralized network, containing preprocessed data from X (formerly Twitter). The data is continuously updated by network miners, providing a real-time stream of tweets for various analytical and machine learning… See the full description on the dataset page: https://huggingface.co/datasets/zephyr-1111/x_dataset_0708150.dclm-260bx_dataset_0701110
Bittensor Subnet 13 X (Twitter) Dataset
Miner Data Compliance Agreement
In uploading this dataset, I am agreeing to the Macrocosmos Miner Data Compliance Policy.
Dataset Summary
This dataset is part of the Bittensor Subnet 13 decentralized network, containing preprocessed data from X (formerly Twitter). The data is continuously updated by network miners, providing a real-time stream of tweets for various analytical and machine learning… See the full description on the dataset page: https://huggingface.co/datasets/zephyr-1111/x_dataset_0701110.details_HuggingFaceH4__zephyr-7b-beta
Dataset Card for Evaluation run of HuggingFaceH4/zephyr-7b-beta
Dataset Summary
Dataset automatically created during the evaluation run of model HuggingFaceH4/zephyr-7b-beta on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_HuggingFaceH4__zephyr-7b-beta.cache-mainx_dataset_070439
Bittensor Subnet 13 X (Twitter) Dataset
Miner Data Compliance Agreement
In uploading this dataset, I am agreeing to the Macrocosmos Miner Data Compliance Policy.
Dataset Summary
This dataset is part of the Bittensor Subnet 13 decentralized network, containing preprocessed data from X (formerly Twitter). The data is continuously updated by network miners, providing a real-time stream of tweets for various analytical and machine learning… See the full description on the dataset page: https://huggingface.co/datasets/zephyr-1111/x_dataset_070439.x_dataset_0707238
Bittensor Subnet 13 X (Twitter) Dataset
Miner Data Compliance Agreement
In uploading this dataset, I am agreeing to the Macrocosmos Miner Data Compliance Policy.
Dataset Summary
This dataset is part of the Bittensor Subnet 13 decentralized network, containing preprocessed data from X (formerly Twitter). The data is continuously updated by network miners, providing a real-time stream of tweets for various analytical and machine learning… See the full description on the dataset page: https://huggingface.co/datasets/zephyr-1111/x_dataset_0707238.x_dataset_0703124
Bittensor Subnet 13 X (Twitter) Dataset
Miner Data Compliance Agreement
In uploading this dataset, I am agreeing to the Macrocosmos Miner Data Compliance Policy.
Dataset Summary
This dataset is part of the Bittensor Subnet 13 decentralized network, containing preprocessed data from X (formerly Twitter). The data is continuously updated by network miners, providing a real-time stream of tweets for various analytical and machine learning… See the full description on the dataset page: https://huggingface.co/datasets/zephyr-1111/x_dataset_0703124.x_dataset_070630
Bittensor Subnet 13 X (Twitter) Dataset
Miner Data Compliance Agreement
In uploading this dataset, I am agreeing to the Macrocosmos Miner Data Compliance Policy.
Dataset Summary
This dataset is part of the Bittensor Subnet 13 decentralized network, containing preprocessed data from X (formerly Twitter). The data is continuously updated by network miners, providing a real-time stream of tweets for various analytical and machine learning… See the full description on the dataset page: https://huggingface.co/datasets/zephyr-1111/x_dataset_070630.cache-128details_BarraHome__zephyr-dpo-v2
Dataset Card for Evaluation run of BarraHome/zephyr-dpo-v2
Dataset automatically created during the evaluation run of model BarraHome/zephyr-dpo-v2 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_BarraHome__zephyr-dpo-v2.x_dataset_0710195
Bittensor Subnet 13 X (Twitter) Dataset
Miner Data Compliance Agreement
In uploading this dataset, I am agreeing to the Macrocosmos Miner Data Compliance Policy.
Dataset Summary
This dataset is part of the Bittensor Subnet 13 decentralized network, containing preprocessed data from X (formerly Twitter). The data is continuously updated by network miners, providing a real-time stream of tweets for various analytical and machine learning… See the full description on the dataset page: https://huggingface.co/datasets/zephyr-1111/x_dataset_0710195.details_alignment-handbook__zephyr-7b-sft-full
Dataset Card for Evaluation run of alignment-handbook/zephyr-7b-sft-full
Dataset automatically created during the evaluation run of model alignment-handbook/zephyr-7b-sft-full on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_alignment-handbook__zephyr-7b-sft-full.details_HuggingFaceH4__zephyr-7b-gemma-v0.1
Dataset Card for Evaluation run of HuggingFaceH4/zephyr-7b-gemma-v0.1
Dataset automatically created during the evaluation run of model HuggingFaceH4/zephyr-7b-gemma-v0.1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_HuggingFaceH4__zephyr-7b-gemma-v0.1.details_UCLA-AGI__zephyr-7b-sft-full-SPIN-iter0
Dataset Card for Evaluation run of UCLA-AGI/zephyr-7b-sft-full-SPIN-iter0
Dataset automatically created during the evaluation run of model UCLA-AGI/zephyr-7b-sft-full-SPIN-iter0 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_UCLA-AGI__zephyr-7b-sft-full-SPIN-iter0.x_dataset_0712117
Bittensor Subnet 13 X (Twitter) Dataset
Miner Data Compliance Agreement
In uploading this dataset, I am agreeing to the Macrocosmos Miner Data Compliance Policy.
Dataset Summary
This dataset is part of the Bittensor Subnet 13 decentralized network, containing preprocessed data from X (formerly Twitter). The data is continuously updated by network miners, providing a real-time stream of tweets for various analytical and machine learning… See the full description on the dataset page: https://huggingface.co/datasets/zephyr-1111/x_dataset_0712117.details_Charlie911__zephyr-7b-beta-MultiLoRA-mmlu-merged
Dataset Card for Evaluation run of Charlie911/zephyr-7b-beta-MultiLoRA-mmlu-merged
Dataset automatically created during the evaluation run of model Charlie911/zephyr-7b-beta-MultiLoRA-mmlu-merged on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Charlie911__zephyr-7b-beta-MultiLoRA-mmlu-merged.Zephyrus
ZephyrusBench
ZephyrusBench is a weather-science benchmark released with the paper Zephyrus: An Agentic Framework for Weather Science. It contains 2,230 question-answer pairs across 49 tasks spanning geospatial reasoning, temporal reasoning, forecasting, simulation, climatology, and scientific question answering.Accepted at the International Conference on Learning Representations, 2026.
Paper and Resources
Paper: arXiv
Poster: ICLR 2026 Poster
Code: Rose-STL-Lab/Zephyrus… See the full description on the dataset page: https://huggingface.co/datasets/Rose-STL-Lab/Zephyrus.honesty_triviaqa_zephyr_responses_v1
Dataset Card for "honesty_zephyr_responses_v1"
More Information needed
details_Charlie911__zephyr-7b-beta-lora-mmlu-merged
Dataset Card for Evaluation run of Charlie911/zephyr-7b-beta-lora-mmlu-merged
Dataset automatically created during the evaluation run of model Charlie911/zephyr-7b-beta-lora-mmlu-merged on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Charlie911__zephyr-7b-beta-lora-mmlu-merged.details_alignment-handbook__zephyr-7b-dpo-full
Dataset Card for Evaluation run of alignment-handbook/zephyr-7b-dpo-full
Dataset automatically created during the evaluation run of model alignment-handbook/zephyr-7b-dpo-full on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_alignment-handbook__zephyr-7b-dpo-full.medmcqa-gen-by-zephyr-ft-gpqa-alldetails_Sao10K__Zephyrus-L1-33B
Dataset Card for Evaluation run of Sao10K/Zephyrus-L1-33B
Dataset Summary
Dataset automatically created during the evaluation run of model Sao10K/Zephyrus-L1-33B on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Sao10K__Zephyrus-L1-33B.details_CallComply__zephyr-7b-beta-128k
Dataset Card for Evaluation run of CallComply/zephyr-7b-beta-128k
Dataset automatically created during the evaluation run of model CallComply/zephyr-7b-beta-128k on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_CallComply__zephyr-7b-beta-128k.details_MexIvanov__zephyr-python-ru
Dataset Card for Evaluation run of MexIvanov/zephyr-python-ru
Dataset automatically created during the evaluation run of model MexIvanov/zephyr-python-ru on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_MexIvanov__zephyr-python-ru.details_dball__zephyr-7b-dpo-qlora-no-sft
Dataset Card for Evaluation run of dball/zephyr-7b-dpo-qlora-no-sft
Dataset automatically created during the evaluation run of model dball/zephyr-7b-dpo-qlora-no-sft on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_dball__zephyr-7b-dpo-qlora-no-sft.details_grimjim__zephyr-wizard-kuno-royale-BF16-merge-7B
Dataset Card for Evaluation run of grimjim/zephyr-wizard-kuno-royale-BF16-merge-7B
Dataset automatically created during the evaluation run of model grimjim/zephyr-wizard-kuno-royale-BF16-merge-7B on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_grimjim__zephyr-wizard-kuno-royale-BF16-merge-7B.
