datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
traces.claude-code.mlx-lm-granitemoehybridibm-granite__granite-3.2-8b-instruct-details
Dataset Card for Evaluation run of ibm-granite/granite-3.2-8b-instruct
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.2-8b-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.2-8b-instruct-details.ibm-granite__granite-3.0-2b-base-details
Dataset Card for Evaluation run of ibm-granite/granite-3.0-2b-base
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.0-2b-base
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.0-2b-base-details.ibm-granite__granite-3.0-2b-instruct-details
Dataset Card for Evaluation run of ibm-granite/granite-3.0-2b-instruct
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.0-2b-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.0-2b-instruct-details.ibm-granite__granite-3.0-1b-a400m-instruct-details
Dataset Card for Evaluation run of ibm-granite/granite-3.0-1b-a400m-instruct
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.0-1b-a400m-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.0-1b-a400m-instruct-details.ibm-granite__granite-3.0-1b-a400m-base-details
Dataset Card for Evaluation run of ibm-granite/granite-3.0-1b-a400m-base
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.0-1b-a400m-base
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.0-1b-a400m-base-details.commercial-tooth-e8f597
commercial-tooth-e8f597
Synthetic weather test data: 53 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/granitetrail/commercial-tooth-e8f597.broad-value-68367a
broad-value-68367a
Synthetic products test data: 59 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/Granite-Rika49/broad-value-68367a.pretty-line-494f34
pretty-line-494f34
Synthetic weather test data: 37 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/graniteVault/pretty-line-494f34.certain-matter-1c98e9
certain-matter-1c98e9
Synthetic products test data: 57 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/GraniteRemy/certain-matter-1c98e9.eastern-card-de9183
eastern-card-de9183
Synthetic weather test data: 43 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/Granite-Minghao/eastern-card-de9183.educational-satisfaction-eafc7f
educational-satisfaction-eafc7f
Synthetic products test data: 42 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number… See the full description on the dataset page: https://huggingface.co/datasets/GraniteEvan/educational-satisfaction-eafc7f.actual-employer-581f8a
actual-employer-581f8a
Synthetic sensors test data: 45 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/GraniteNoah/actual-employer-581f8a.zelo-scores-10kx100-granite-4.1-30b
Dataset Card for tomaarsen/zelo-scores-10kx100-granite-4.1-30b
Dataset Summary
Synthetic data generated by DataForge:
Model: ibm-granite/granite-4.1-30b (main)
Source dataset: tomaarsen/zelo-pairs-10kx100-quantile-anchor (train split).
Generation config: temperature=None, top_p=None, top_k=None, max_tokens=4096, model_max_context=32768
Speculative decoding: disabled
System prompt: `You are a relevance scoring system. Given a query and two documents (A and B), your job… See the full description on the dataset page: https://huggingface.co/datasets/tomaarsen/zelo-scores-10kx100-granite-4.1-30b.potential-peak-de248c
potential-peak-de248c
Synthetic weather test data: 53 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/GraniteStar/potential-peak-de248c.wide-eye-b24176
wide-eye-b24176
Synthetic products test data: 60 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/graniteGlade/wide-eye-b24176.major-disease-a8768f
major-disease-a8768f
Synthetic sensors test data: 39 rows in data.csv.
All values are randomly generated fictional examples, not real observations, products, or user activity. Intended only for CSV loading and pipeline tests; not suitable for scientific or business conclusions. Columns are sampled independently and do not model real-world correlations.
Fields
sample_id: random identifier for this generated sample.
row_id: sequential row number starting at 1.… See the full description on the dataset page: https://huggingface.co/datasets/graniteP/major-disease-a8768f.ibm-granite__granite-3.1-8b-instruct-details
Dataset Card for Evaluation run of ibm-granite/granite-3.1-8b-instruct
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.1-8b-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.1-8b-instruct-details.ibm-granite__granite-7b-base-details
Dataset Card for Evaluation run of ibm-granite/granite-7b-base
Dataset automatically created during the evaluation run of model ibm-granite/granite-7b-base
The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-7b-base-details.ibm-granite__granite-3.2-2b-instruct-details
Dataset Card for Evaluation run of ibm-granite/granite-3.2-2b-instruct
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.2-2b-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.2-2b-instruct-details.ibm-granite__granite-3.0-8b-instruct-details
Dataset Card for Evaluation run of ibm-granite/granite-3.0-8b-instruct
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.0-8b-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.0-8b-instruct-details.ibm-granite__granite-3.1-3b-a800m-base-details
Dataset Card for Evaluation run of ibm-granite/granite-3.1-3b-a800m-base
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.1-3b-a800m-base
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.1-3b-a800m-base-details.ibm-granite__granite-3.0-3b-a800m-instructibm-granite__granite-7b-instruct-details
Dataset Card for Evaluation run of ibm-granite/granite-7b-instruct
Dataset automatically created during the evaluation run of model ibm-granite/granite-7b-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-7b-instruct-details.ibm-granite__granite-3.0-3b-a800m-instruct-details
Dataset Card for Evaluation run of ibm-granite/granite-3.0-3b-a800m-instruct
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.0-3b-a800m-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.0-3b-a800m-instruct-details.ibm-granite__granite-3.0-8b-base-details
Dataset Card for Evaluation run of ibm-granite/granite-3.0-8b-base
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.0-8b-base
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.0-8b-base-details.ibm-granite__granite-3.1-2b-base-details
Dataset Card for Evaluation run of ibm-granite/granite-3.1-2b-base
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.1-2b-base
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.1-2b-base-details.ibm-granite__granite-3.1-8b-base-details
Dataset Card for Evaluation run of ibm-granite/granite-3.1-8b-base
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.1-8b-base
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.1-8b-base-details.ibm-granite__granite-3.1-3b-a800m-instruct-details
Dataset Card for Evaluation run of ibm-granite/granite-3.1-3b-a800m-instruct
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.1-3b-a800m-instruct
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.1-3b-a800m-instruct-details.ibm-granite__granite-3.0-3b-a800m-base-details
Dataset Card for Evaluation run of ibm-granite/granite-3.0-3b-a800m-base
Dataset automatically created during the evaluation run of model ibm-granite/granite-3.0-3b-a800m-base
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ibm-granite__granite-3.0-3b-a800m-base-details.
