datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
homer_math_v0.1homer_math_v0.1 is a dataset that is cleaned from OpenMathInstruct-2 and removes samples similar to the MATH benchmark.
homeroom-copilot-open-traces
Homeroom Copilot Open Traces
This dataset contains sanitized JSONL development trace excerpts from Homeroom Copilot, a teacher-facing educational AI dashboard created for the Build Small Hackathon.
Homeroom Copilot combines deterministic student risk assessment, root-cause analysis, curated evidence-based intervention retrieval, and AI-assisted action-plan generation for middle school teachers. These traces document selected Codex-assisted development moments from the project.… See the full description on the dataset page: https://huggingface.co/datasets/ravi2505/homeroom-copilot-open-traces.mn-context-compression-dataset-v1
MN Context Compression Dataset v1
Author: Homer Quan
This dataset is used to train context-compression models for improving the context efficiency of multi-agent runtimes, especially MirrorNeuron and the broader work at mirrorneuron.io.
We use this dataset to train models such as homerquan/mn-context-engine-lora-v2, and later protected-fact-focused context engines. The data emphasizes exact protected-span retention, source-reference preservation, budget-conditioned compression, and… See the full description on the dataset page: https://huggingface.co/datasets/homerquan/mn-context-compression-dataset-v1.allknowingroger__HomerSlerp2-7B-details
Dataset Card for Evaluation run of allknowingroger/HomerSlerp2-7B
Dataset automatically created during the evaluation run of model allknowingroger/HomerSlerp2-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allknowingroger__HomerSlerp2-7B-details.newsbang__Homer-v1.0-Qwen2.5-72B-details
Dataset Card for Evaluation run of newsbang/Homer-v1.0-Qwen2.5-72B
Dataset automatically created during the evaluation run of model newsbang/Homer-v1.0-Qwen2.5-72B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-v1.0-Qwen2.5-72B-details.newsbang__Homer-v0.4-Qwen2.5-7B-details
Dataset Card for Evaluation run of newsbang/Homer-v0.4-Qwen2.5-7B
Dataset automatically created during the evaluation run of model newsbang/Homer-v0.4-Qwen2.5-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-v0.4-Qwen2.5-7B-details.boardgamebench-answer-dpo
BoardGameBench Answer DPO Dataset
This dataset contains the reviewed preference examples used for the DPO stage of the nemotron-boardgame-answer-lora-b4-safe-final adapter.
It is a compact pilot set of 10 BoardGameBench preference rows. Each row presents the same board-game decision prompt with a preferred answer and a plausible rejected answer. The preferred answer is selected from engine-guided move comparisons and includes the exact move label.
Format
The main… See the full description on the dataset page: https://huggingface.co/datasets/homerquan/boardgamebench-answer-dpo.hotmailuser__Qwen2.5-HomerSlerp-7B-details
Dataset Card for Evaluation run of hotmailuser/Qwen2.5-HomerSlerp-7B
Dataset automatically created during the evaluation run of model hotmailuser/Qwen2.5-HomerSlerp-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/hotmailuser__Qwen2.5-HomerSlerp-7B-details.ZeroXClem__Qwen2.5-7B-HomerCreative-Mix-details
Dataset Card for Evaluation run of ZeroXClem/Qwen2.5-7B-HomerCreative-Mix
Dataset automatically created during the evaluation run of model ZeroXClem/Qwen2.5-7B-HomerCreative-Mix
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ZeroXClem__Qwen2.5-7B-HomerCreative-Mix-details.ZeroXClem__Qwen2.5-7B-HomerAnvita-NerdMix-details
Dataset Card for Evaluation run of ZeroXClem/Qwen2.5-7B-HomerAnvita-NerdMix
Dataset automatically created during the evaluation run of model ZeroXClem/Qwen2.5-7B-HomerAnvita-NerdMix
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/ZeroXClem__Qwen2.5-7B-HomerAnvita-NerdMix-details.jiangxinyang-shanda__Homer-LLama3-8B-details
Dataset Card for Evaluation run of jiangxinyang-shanda/Homer-LLama3-8B
Dataset automatically created during the evaluation run of model jiangxinyang-shanda/Homer-LLama3-8B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/jiangxinyang-shanda__Homer-LLama3-8B-details.newsbang__Homer-7B-v0.1-details
Dataset Card for Evaluation run of newsbang/Homer-7B-v0.1
Dataset automatically created during the evaluation run of model newsbang/Homer-7B-v0.1
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-7B-v0.1-details.suayptalha__HomerCreativeAnvita-Mix-Qw7B-details
Dataset Card for Evaluation run of suayptalha/HomerCreativeAnvita-Mix-Qw7B
Dataset automatically created during the evaluation run of model suayptalha/HomerCreativeAnvita-Mix-Qw7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/suayptalha__HomerCreativeAnvita-Mix-Qw7B-details.HomerSimpsonsFullnewsbang__Homer-v1.0-Qwen2.5-7B-details
Dataset Card for Evaluation run of newsbang/Homer-v1.0-Qwen2.5-7B
Dataset automatically created during the evaluation run of model newsbang/Homer-v1.0-Qwen2.5-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-v1.0-Qwen2.5-7B-details.HomerSimpsonnewsbang__Homer-v0.3-Qwen2.5-7B-details
Dataset Card for Evaluation run of newsbang/Homer-v0.3-Qwen2.5-7B
Dataset automatically created during the evaluation run of model newsbang/Homer-v0.3-Qwen2.5-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-v0.3-Qwen2.5-7B-details.allknowingroger__HomerSlerp4-7B-details
Dataset Card for Evaluation run of allknowingroger/HomerSlerp4-7B
Dataset automatically created during the evaluation run of model allknowingroger/HomerSlerp4-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allknowingroger__HomerSlerp4-7B-details.newsbang__Homer-7B-v0.2-details
Dataset Card for Evaluation run of newsbang/Homer-7B-v0.2
Dataset automatically created during the evaluation run of model newsbang/Homer-7B-v0.2
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-7B-v0.2-details.newsbang__Homer-v0.5-Qwen2.5-7B-details
Dataset Card for Evaluation run of newsbang/Homer-v0.5-Qwen2.5-7B
Dataset automatically created during the evaluation run of model newsbang/Homer-v0.5-Qwen2.5-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/newsbang__Homer-v0.5-Qwen2.5-7B-details.allknowingroger__HomerSlerp1-7B-details
Dataset Card for Evaluation run of allknowingroger/HomerSlerp1-7B
Dataset automatically created during the evaluation run of model allknowingroger/HomerSlerp1-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allknowingroger__HomerSlerp1-7B-details.allknowingroger__HomerSlerp3-7B-details
Dataset Card for Evaluation run of allknowingroger/HomerSlerp3-7B
Dataset automatically created during the evaluation run of model allknowingroger/HomerSlerp3-7B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/allknowingroger__HomerSlerp3-7B-details.homerobot_01
