datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
git-history-mcq-ru
git-history-mcq-ru
805 вопросов с вариантами ответа по истории трёх открытых репозиториев
(digitable-lol/digit, digitable-lol/digitwm, digitable-lol/flang), плюс
8 672 ответа пяти моделей и 4 878 разборов этих ответов.
Вопросы на русском. Ключ каждого выведен из вывода git-команды, и сама команда
и её вывод лежат в записи — задачу можно перепроверить, не доверяя составителю.
Набор собран для одной проверки: меняют ли что-нибудь приёмы промптинга. Девять
вариантов оформления… See the full description on the dataset page: https://huggingface.co/datasets/the-homeless-god/git-history-mcq-ru.homeroom-copilot-open-traces
Homeroom Copilot Open Traces
This dataset contains sanitized JSONL development trace excerpts from Homeroom Copilot, a teacher-facing educational AI dashboard created for the Build Small Hackathon.
Homeroom Copilot combines deterministic student risk assessment, root-cause analysis, curated evidence-based intervention retrieval, and AI-assisted action-plan generation for middle school teachers. These traces document selected Codex-assisted development moments from the project.… See the full description on the dataset page: https://huggingface.co/datasets/ravi2505/homeroom-copilot-open-traces.allknowingroger__HomerSlerp2-7B-details
Dataset Card for Evaluation run of allknowingroger/HomerSlerp2-7B
Dataset automatically created during the evaluation run of model allknowingroger/HomerSlerp2-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allknowingroger__HomerSlerp2-7B-details.newsbang__Homer-v1.0-Qwen2.5-72B-details
Dataset Card for Evaluation run of newsbang/Homer-v1.0-Qwen2.5-72B
Dataset automatically created during the evaluation run of model newsbang/Homer-v1.0-Qwen2.5-72B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-v1.0-Qwen2.5-72B-details.digitable-cluster-cells
Ячейки кластерной работы: бриф → прогон → исход
30 записей о работе кластера ИИ-агентов над тремя открытыми репозиториями
(digitwm, dotfiles, digit) 30–31 августа 2026. Одна запись — одна ячейка
работы: что поручили, каким брифом, что прогнали, какие числа получили и чем
кончилось.
Набор собран не ради демонстрации успехов. Он существует, чтобы утверждение
«подробный бриф и кластерное устройство дают лучший результат» можно было
опровергнуть, а не только проиллюстрировать.… See the full description on the dataset page: https://huggingface.co/datasets/the-homeless-god/digitable-cluster-cells.newsbang__Homer-v0.4-Qwen2.5-7B-details
Dataset Card for Evaluation run of newsbang/Homer-v0.4-Qwen2.5-7B
Dataset automatically created during the evaluation run of model newsbang/Homer-v0.4-Qwen2.5-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-v0.4-Qwen2.5-7B-details.hotmailuser__Qwen2.5-HomerSlerp-7B-details
Dataset Card for Evaluation run of hotmailuser/Qwen2.5-HomerSlerp-7B
Dataset automatically created during the evaluation run of model hotmailuser/Qwen2.5-HomerSlerp-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/hotmailuser__Qwen2.5-HomerSlerp-7B-details.ZeroXClem__Qwen2.5-7B-HomerCreative-Mix-details
Dataset Card for Evaluation run of ZeroXClem/Qwen2.5-7B-HomerCreative-Mix
Dataset automatically created during the evaluation run of model ZeroXClem/Qwen2.5-7B-HomerCreative-Mix
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ZeroXClem__Qwen2.5-7B-HomerCreative-Mix-details.ZeroXClem__Qwen2.5-7B-HomerAnvita-NerdMix-details
Dataset Card for Evaluation run of ZeroXClem/Qwen2.5-7B-HomerAnvita-NerdMix
Dataset automatically created during the evaluation run of model ZeroXClem/Qwen2.5-7B-HomerAnvita-NerdMix
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ZeroXClem__Qwen2.5-7B-HomerAnvita-NerdMix-details.jiangxinyang-shanda__Homer-LLama3-8B-details
Dataset Card for Evaluation run of jiangxinyang-shanda/Homer-LLama3-8B
Dataset automatically created during the evaluation run of model jiangxinyang-shanda/Homer-LLama3-8B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/jiangxinyang-shanda__Homer-LLama3-8B-details.newsbang__Homer-7B-v0.1-details
Dataset Card for Evaluation run of newsbang/Homer-7B-v0.1
Dataset automatically created during the evaluation run of model newsbang/Homer-7B-v0.1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-7B-v0.1-details.suayptalha__HomerCreativeAnvita-Mix-Qw7B-details
Dataset Card for Evaluation run of suayptalha/HomerCreativeAnvita-Mix-Qw7B
Dataset automatically created during the evaluation run of model suayptalha/HomerCreativeAnvita-Mix-Qw7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/suayptalha__HomerCreativeAnvita-Mix-Qw7B-details.forager-sightings
Forager Sightings
A growing, crowd-contributed dataset of wild plant, mushroom, and berry photos
with human labels, collected through the "Beat the Machine" game in
Forager's Field Station.
Every row pairs a field photo with the contributor's own identification and
the on-device model's prediction (or its refusal). It is an open-research dataset
for improving small, on-device foraging models — especially the hard cases the
model abstains on.
How it's collected… See the full description on the dataset page: https://huggingface.co/datasets/HomesteaderLabs/forager-sightings.newsbang__Homer-v1.0-Qwen2.5-7B-details
Dataset Card for Evaluation run of newsbang/Homer-v1.0-Qwen2.5-7B
Dataset automatically created during the evaluation run of model newsbang/Homer-v1.0-Qwen2.5-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-v1.0-Qwen2.5-7B-details.low-high-dir-20260728-home-l-liyj-shiying-KIVI
Ouzhang/low-high-dir-20260728-home-l-liyj-shiying-KIVI
Source directory: /home/l/liyj/shiying/KIVI
Files: 2537
Bytes: 113008623190
Uploaded with upload_large_folder on 2026-07-28.
newsbang__Homer-v0.3-Qwen2.5-7B-details
Dataset Card for Evaluation run of newsbang/Homer-v0.3-Qwen2.5-7B
Dataset automatically created during the evaluation run of model newsbang/Homer-v0.3-Qwen2.5-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-v0.3-Qwen2.5-7B-details.allknowingroger__HomerSlerp4-7B-details
Dataset Card for Evaluation run of allknowingroger/HomerSlerp4-7B
Dataset automatically created during the evaluation run of model allknowingroger/HomerSlerp4-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allknowingroger__HomerSlerp4-7B-details.newsbang__Homer-7B-v0.2-details
Dataset Card for Evaluation run of newsbang/Homer-7B-v0.2
Dataset automatically created during the evaluation run of model newsbang/Homer-7B-v0.2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-7B-v0.2-details.newsbang__Homer-v0.5-Qwen2.5-7B-details
Dataset Card for Evaluation run of newsbang/Homer-v0.5-Qwen2.5-7B
Dataset automatically created during the evaluation run of model newsbang/Homer-v0.5-Qwen2.5-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-v0.5-Qwen2.5-7B-details.allknowingroger__HomerSlerp1-7B-details
Dataset Card for Evaluation run of allknowingroger/HomerSlerp1-7B
Dataset automatically created during the evaluation run of model allknowingroger/HomerSlerp1-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allknowingroger__HomerSlerp1-7B-details.allknowingroger__HomerSlerp3-7B-details
Dataset Card for Evaluation run of allknowingroger/HomerSlerp3-7B
Dataset automatically created during the evaluation run of model allknowingroger/HomerSlerp3-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allknowingroger__HomerSlerp3-7B-details.
