homer
Datasets
All datasets matching “homer”alotb_v0This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "panda",
"total_episodes": 300,
"total_frames": 50521,
"total_tasks": 3,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:300"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path": "videos/{video_key}/chunk-{chunk_index:03d}/file-{file_index:03d}.mp4",
"features": {… See the full description on the dataset page: https://huggingface.co/datasets/HomeRobotics/alotb_v0.boardgamebench-answer-only
BoardGameBench Answer-Only Reasoning Dataset
This dataset contains 1,282,766 board-game reasoning examples generated from BoardGameBench, a benchmark and data-generation project for evaluating language models on structured board-game decision making.
Each row asks a model to inspect a legal board position and return the best move. The format is intentionally simple:
id,prompt,answer
The answer field is the target move label, such as C4, f6, 11,7, or e2-d3. This makes the dataset… See the full description on the dataset page: https://huggingface.co/datasets/homerquan/boardgamebench-answer-only.homer-v2
HomER v2: Home Egocentric Robotics Dataset
HomER v2 is a curated dataset of egocentric household activity videos designed for robotics, embodied AI, and video understanding research.
The dataset contains 765 first-person videos representing approximately 100 hours of real-world household activities collected from 420 unique participants across 49 countries spanning 5 continents. Each video is accompanied by structured metadata and a natural-language description of the performed… See the full description on the dataset page: https://huggingface.co/datasets/toloka/homer-v2.homer_math_v0.1homer_math_v0.1 is a dataset that is cleaned from OpenMathInstruct-2 and removes samples similar to the MATH benchmark.
boardgamebench-answer-sft
BoardGameBench Answer SFT Dataset
This dataset contains 1,282,766 answer-only board-game reasoning examples generated from BoardGameBench, a benchmark and data-generation project for evaluating language models on structured board-game decision making.
This is the supervised fine-tuning corpus used before the later DPO and GRPO stages for the nemotron-boardgame-answer-lora-b4-safe-final adapter.
Each row asks a model to inspect a legal board position and return the best move. The… See the full description on the dataset page: https://huggingface.co/datasets/homerquan/boardgamebench-answer-sft.homeroom-copilot-open-traces
Homeroom Copilot Open Traces
This dataset contains sanitized JSONL development trace excerpts from Homeroom Copilot, a teacher-facing educational AI dashboard created for the Build Small Hackathon.
Homeroom Copilot combines deterministic student risk assessment, root-cause analysis, curated evidence-based intervention retrieval, and AI-assisted action-plan generation for middle school teachers. These traces document selected Codex-assisted development moments from the project.… See the full description on the dataset page: https://huggingface.co/datasets/ravi2505/homeroom-copilot-open-traces.
