datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
officeqa-checkpoint-eval-data
Checkpoint evaluation plot data
Snapshot: 2026-09-14T16:26:45.684890+00:00. Aggregate inputs to notes/Sept-2-2026.md performance figures.
No model execution, grading, publication, or source-result changes were performed to make this export.
Contents
checkpoint_evaluations: 454 checkpoint rows, one evaluation per run/iteration/protocol; score, mean output tokens, mean steps, and the existing two-sided 95% confidence bounds.
pareto_points: current mean-token/USD… See the full description on the dataset page: https://huggingface.co/datasets/YWZBrandon/officeqa-checkpoint-eval-data.DICE-BENCH
🎲 DICE-BENCH: Evaluating the Tool-Use Capabilities of Large Language Models in Multi-Round, Multi-Party Dialogues
🔗 Links for Reference
Repository: https://github.com/snuhcc/DICE-Bench
Paper: https://arxiv.org/abs/2506.22853
Project page: https://snuhcc.github.io/DICE-Bench/
Point of Contact: kyochul@snu.ac.kr
📖 Paper Description
DICE-BENCH is a benchmark that tests how well large language models can call external functions in realistic… See the full description on the dataset page: https://huggingface.co/datasets/OfficerChul/DICE-BENCH.OfficeFatigue
OfficeFatigue Signal-Only Release
This initial release contains participant-level processed signal arrays for OfficeFatigue. It includes 25 participants (p01-p25) and two subsets per participant:
OFS_signal.npy: short controlled desk-work subset.
OFL_signal.npy: long natural office subset.
Fatigue labels, benchmark splits, and diagnostic/oracle-corrected labels are not included in this signal-only release. Labels will be released separately in a later version.
File… See the full description on the dataset page: https://huggingface.co/datasets/Officefatigue/OfficeFatigue.OfficeSmith-PPTX-IR
OfficeSmith PPTX IR
Synthetic bilingual business briefs paired with editable PPTX intermediate representations.
Dataset summary
This dataset is part of the OfficeSmith collection for training models to plan, build, clarify, critique, and repair editable business presentations. It contains observable outputs only: no hidden chain of thought, secret benchmark prompt, personal data, or API credential is included.
Train rows: 862
Validation rows: 48
Test rows: 48… See the full description on the dataset page: https://huggingface.co/datasets/Benitoow/OfficeSmith-PPTX-IR.govie-office-holder-reranker-bilingual-v2
gov.ie Office Holder Reranker Bilingual v2
Bilingual (query, candidate_page) reranking dataset for current Irish government office-holder lookup on public gov.ie pages.
This is a derivative of temsa/govie-office-holder-reranker-dataset-v1 with one important change: every English query is paired with an Irish Gaelic query variant while keeping the same candidate pool and labels.
What changed versus v1
same candidate pages and labels
same office-holder snapshot and… See the full description on the dataset page: https://huggingface.co/datasets/temsa/govie-office-holder-reranker-bilingual-v2.OfficeSmith-PPTX-Repair
OfficeSmith PPTX Repair
Deterministically degraded PPTX IR objects paired with validated repairs.
Dataset summary
This dataset is part of the OfficeSmith collection for training models to plan, build, clarify, critique, and repair editable business presentations. It contains observable outputs only: no hidden chain of thought, secret benchmark prompt, personal data, or API credential is included.
Train rows: 160
Validation rows: 0
Test rows: 0
Languages: French… See the full description on the dataset page: https://huggingface.co/datasets/Benitoow/OfficeSmith-PPTX-Repair.Android-Control-84k
Android Control Dataset
Overview
This directory contains two dataset files (and_ctrl_train.json and and_ctrl_test.json) derived from the Android Control project by Google Research. These datasets have been formatted specifically for GUI grounding training in LLaMA-Factory.
Dataset Description
The Android Control dataset consists of episodes where each episode contains multiple steps. Each step includes:
Step instructions: Natural language instructions for UI… See the full description on the dataset page: https://huggingface.co/datasets/OfficerChul/Android-Control-84k.OfficeSmith-PPTX-Clarify-Critique
OfficeSmith PPTX Clarification and Critique
Ambiguous presentation requests and structured PPTX IR critiques.
Dataset summary
This dataset is part of the OfficeSmith collection for training models to plan, build, clarify, critique, and repair editable business presentations. It contains observable outputs only: no hidden chain of thought, secret benchmark prompt, personal data, or API credential is included.
Train rows: 160
Validation rows: 0
Test rows: 0… See the full description on the dataset page: https://huggingface.co/datasets/Benitoow/OfficeSmith-PPTX-Clarify-Critique.office-holder-policy-reranker-v1
Office Holder Policy Reranker v1
Evaluation and training dataset for government office-holder reranking under a serving-like candidate representation.
Purpose:
current office-holder identity queries such as Who is the Taoiseach?
English and Irish Gaelic role-identity queries
candidate documents rendered in short structured form closer to serving input:
Title:
Category:
Organisation:
URL:
Summary: from the opening sentence(s)
hard negatives including:
former office-holder pages… See the full description on the dataset page: https://huggingface.co/datasets/temsa/office-holder-policy-reranker-v1.the_office_finetome100k_format.jsonl
the_office_finetome100k_format.jsonl
Files
train.jsonl
Loading (example)
from datasets import load_dataset
ds = load_dataset("Mathieu-Thomas-JOSSET/the_office_finetome100k_format.jsonl")
print(ds)
govie-office-holder-reranker-dataset-v1
gov.ie Office Holder Reranker Dataset v1
Binary reranking dataset for current Irish government office-holder lookup on public gov.ie pages.
This is not a directory of office holders by itself. It is a ranking dataset where each row is a (query, candidate_page) pair labeled as relevant (1) or not relevant (0).
For users who want the compact canonical office-holder list behind the training pairs, see metadata/office_holders.json.
What A Row Means
Each row contains:… See the full description on the dataset page: https://huggingface.co/datasets/temsa/govie-office-holder-reranker-dataset-v1.PS_AD_Office365_05_ShareGPTthe_office_only_michael_finetome_no_le2w_fullnames_v2.jsonl
the_office_only_michael_finetome_no_le2w_fullnames_v2.jsonl
Files
train.jsonl
Loading (example)
from datasets import load_dataset
ds = load_dataset("Mathieu-Thomas-JOSSET/the_office_only_michael_finetome_no_le2w_fullnames_v2.jsonl")
print(ds)
linhhuonglinux-office-dataset-v3
🚀 Linh Hương Linux Office Dataset V3 (Master/Production Ready)
Đây là bộ dữ liệu khổng lồ thế hệ mới nhất dành riêng cho việc huấn luyện Linh Hương Linux AI.
Tập dữ liệu này được chia làm 2 cấu hình (Configs) riêng biệt để tránh lỗi xung đột cấu trúc (Schema Conflict):
1. Cấu hình SFT (sft)
Chứa 1,783 tình huống Supervised Fine-Tuning, làm sạch 100% bằng Heuristic Rejection Sampling.
Multi-turn Chat: Hội thoại nhiều lượt (nhớ bối cảnh).
Function Calling: Gọi hàm điều… See the full description on the dataset page: https://huggingface.co/datasets/linhhuonglinux/linhhuonglinux-office-dataset-v3.humanoid-office-assistant-tr-v1Office assistant interaction dataset.
Description
Basic office support and task execution interactions for humanoid robots.
Task Description
Teaches humanoid robots to assist in office environments by preparing meeting rooms, managing schedules and generating daily operational reports.
the_office_only_michael.jsonl
the_office_only_michael.jsonl
Files
train.jsonl
Loading (example)
from datasets import load_dataset
ds = load_dataset("Mathieu-Thomas-JOSSET/the_office_only_michael.jsonl")
print(ds)
Synthetic-Office-SupplyPS_AD_Office365_02Second version of the synthetic dataset created by putting a part of a textbook in the context of 7B model and then asking the model
to create a few questions and answers related to the dataset.
It contains information about PowerShell basics, Office 365 basics and Active Directory/GPO basics.
PS_AD_Office365_03Previous version with a subset of spicyboros 2.2 coding samples plus some a few other new PowerShell scripting samples. Some formatting fixes.
PS_AD_Office_01Synthetic dataset of PowerShell, Active Directory and I think some Office 365 Q&A
PS_AD_Office365_04_ShareGPTthe_office_only_michael_finetome_no_le2w.jsonl
the_office_only_michael_finetome_no_le2w.jsonl
Files
train.jsonl
Loading (example)
from datasets import load_dataset
ds = load_dataset("Mathieu-Thomas-JOSSET/the_office_only_michael_finetome_no_le2w.jsonl")
print(ds)
the_office_only_michael_finetome_.jsonl
the_office_only_michael_finetome_.jsonl
Files
train.jsonl
Loading (example)
from datasets import load_dataset
ds = load_dataset("Mathieu-Thomas-JOSSET/the_office_only_michael_finetome_.jsonl")
print(ds)
office1office-translations-1k
