datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
open_lm_test_data_v2crystallography-open-database
Crystallography Open Database (COD) — Full Snapshot
A complete mirror of the Crystallography Open Database (COD) as a single Parquet file, combining all crystallographic metadata with the raw CIF file content in one queryable dataset.
Snapshot Details
Field
Value
Snapshot date
2026-07-06
Metadata fetched
2026-07-06 18:51 (UTC+2) — 533,486 entries
CIF files downloaded
2026-07-06 18:34–21:58 — 533,862 files
Total rows
533,486 (metadata) — 411… See the full description on the dataset page: https://huggingface.co/datasets/LMucko/crystallography-open-database.details_openlm-research__open_llama_3b
Dataset Card for Evaluation run of openlm-research/open_llama_3b
Dataset Summary
Dataset automatically created during the evaluation run of model openlm-research/open_llama_3b on the Open LLM Leaderboard.
The dataset is composed of 122 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 6 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_openlm-research__open_llama_3b.open_lm_example_datamultimodal-open-r1-8k-verifieddetails_openlm-research__open_llama_7b_v2
Dataset Card for Evaluation run of openlm-research/open_llama_7b_v2
Dataset Summary
Dataset automatically created during the evaluation run of model openlm-research/open_llama_7b_v2 on the Open LLM Leaderboard.
The dataset is composed of 122 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_openlm-research__open_llama_7b_v2.details_openlm-research__open_llama_7b
Dataset Card for Evaluation run of openlm-research/open_llama_7b
Dataset Summary
Dataset automatically created during the evaluation run of model openlm-research/open_llama_7b on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_openlm-research__open_llama_7b.details_Yhyu13__LMCocktail-10.7B-v1
Dataset Card for Evaluation run of yhyu13/LMCocktail-10.7B-v1
Dataset automatically created during the evaluation run of model yhyu13/LMCocktail-10.7B-v1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Yhyu13__LMCocktail-10.7B-v1.details_openlm-research__open_llama_13b
Dataset Card for Evaluation run of openlm-research/open_llama_13b
Dataset Summary
Dataset automatically created during the evaluation run of model openlm-research/open_llama_13b on the Open LLM Leaderboard.
The dataset is composed of 122 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 5 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_openlm-research__open_llama_13b.open_lm_test_datadetails_lmsys__vicuna-7b-v1.3
Dataset Card for Evaluation run of lmsys/vicuna-7b-v1.3
Dataset Summary
Dataset automatically created during the evaluation run of model lmsys/vicuna-7b-v1.3 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_lmsys__vicuna-7b-v1.3.open-lm-instruction-datadetails_Xwin-LM__Xwin-Math-70B-V1.0
Dataset Card for Evaluation run of Xwin-LM/Xwin-Math-70B-V1.0
Dataset automatically created during the evaluation run of model Xwin-LM/Xwin-Math-70B-V1.0 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Xwin-LM__Xwin-Math-70B-V1.0.details_invalid-coder__Starling-LM-7B-beta-laser-dpo
Dataset Card for Evaluation run of invalid-coder/Starling-LM-7B-beta-laser-dpo
Dataset automatically created during the evaluation run of model invalid-coder/Starling-LM-7B-beta-laser-dpo on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_invalid-coder__Starling-LM-7B-beta-laser-dpo.details_Yhyu13__LMCocktail-Mistral-7B-v1
Dataset Card for Evaluation run of Yhyu13/LMCocktail-Mistral-7B-v1
Dataset automatically created during the evaluation run of model Yhyu13/LMCocktail-Mistral-7B-v1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Yhyu13__LMCocktail-Mistral-7B-v1.details_Xwin-LM__Xwin-LM-70B-V0.1
Dataset Card for Evaluation run of Xwin-LM/Xwin-LM-70B-V0.1
Dataset Summary
Dataset automatically created during the evaluation run of model Xwin-LM/Xwin-LM-70B-V0.1 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Xwin-LM__Xwin-LM-70B-V0.1.details_karakuri-ai__karakuri-lm-70b-chat-v0.1
Dataset Card for Evaluation run of karakuri-ai/karakuri-lm-70b-chat-v0.1
Dataset automatically created during the evaluation run of model karakuri-ai/karakuri-lm-70b-chat-v0.1 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_karakuri-ai__karakuri-lm-70b-chat-v0.1.details_Xwin-LM__Xwin-LM-7B-V0.1
Dataset Card for Evaluation run of Xwin-LM/Xwin-LM-7B-V0.1
Dataset Summary
Dataset automatically created during the evaluation run of model Xwin-LM/Xwin-LM-7B-V0.1 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Xwin-LM__Xwin-LM-7B-V0.1.details_lmsys__vicuna-7b-v1.5
Dataset Card for Evaluation run of lmsys/vicuna-7b-v1.5
Dataset Summary
Dataset automatically created during the evaluation run of model lmsys/vicuna-7b-v1.5 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_lmsys__vicuna-7b-v1.5.details_lmsys__vicuna-13b-v1.5-16k
Dataset Card for Evaluation run of lmsys/vicuna-13b-v1.5-16k
Dataset Summary
Dataset automatically created during the evaluation run of model lmsys/vicuna-13b-v1.5-16k on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_lmsys__vicuna-13b-v1.5-16k.details_Community-LM__llava-v1.5-13b-hf
Dataset Card for Evaluation run of Community-LM/llava-v1.5-13b-hf
Dataset Summary
Dataset automatically created during the evaluation run of model Community-LM/llava-v1.5-13b-hf on the Open LLM Leaderboard.
The dataset is composed of 61 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_Community-LM__llava-v1.5-13b-hf.details_lmsys__vicuna-7b-v1.5-16k
Dataset Card for Evaluation run of lmsys/vicuna-7b-v1.5-16k
Dataset Summary
Dataset automatically created during the evaluation run of model lmsys/vicuna-7b-v1.5-16k on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 3 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_lmsys__vicuna-7b-v1.5-16k.details_lmsys__vicuna-13b-v1.5
Dataset Card for Evaluation run of lmsys/vicuna-13b-v1.5
Dataset Summary
Dataset automatically created during the evaluation run of model lmsys/vicuna-13b-v1.5 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_lmsys__vicuna-13b-v1.5.details_lmsys__vicuna-13b-v1.1
Dataset Card for Evaluation run of lmsys/vicuna-13b-v1.1
Dataset Summary
Dataset automatically created during the evaluation run of model lmsys/vicuna-13b-v1.1 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_lmsys__vicuna-13b-v1.1.details_CallComply__Starling-LM-11B-alpha
Dataset Card for Evaluation run of CallComply/Starling-LM-11B-alpha
Dataset automatically created during the evaluation run of model CallComply/Starling-LM-11B-alpha on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_CallComply__Starling-LM-11B-alpha.details_perlthoughts__Starling-LM-alpha-8x7B-MoE
Dataset Card for Evaluation run of perlthoughts/Starling-LM-alpha-8x7B-MoE
Dataset automatically created during the evaluation run of model perlthoughts/Starling-LM-alpha-8x7B-MoE on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_perlthoughts__Starling-LM-alpha-8x7B-MoE.details_TeeZee__Xwin-LM-70B-V0.1_Limarpv3
Dataset Card for Evaluation run of TeeZee/Xwin-LM-70B-V0.1_Limarpv3
Dataset automatically created during the evaluation run of model TeeZee/Xwin-LM-70B-V0.1_Limarpv3 on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_TeeZee__Xwin-LM-70B-V0.1_Limarpv3.Openlm1details_lmsys__vicuna-13b-delta-v1.1
Dataset Card for Evaluation run of lmsys/vicuna-13b-delta-v1.1
Dataset Summary
Dataset automatically created during the evaluation run of model lmsys/vicuna-13b-delta-v1.1 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_lmsys__vicuna-13b-delta-v1.1.details_lmsys__vicuna-33b-v1.3
Dataset Card for Evaluation run of lmsys/vicuna-33b-v1.3
Dataset Summary
Dataset automatically created during the evaluation run of model lmsys/vicuna-33b-v1.3 on the Open LLM Leaderboard.
The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_lmsys__vicuna-33b-v1.3.
