datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
fish_datasets_real_electrodyn_expertsys_twodim_fourierrobopro_expertgpt-oss-20b-moe-expert-power-traces-320k
GPT-OSS-20B MoE Expert Power Traces (320k, ChipWhisperer)
This dataset contains analog power traces captured with a ChipWhisperer Husky while running forced single-expert MoE computations derived from openai/gpt-oss-20b on an NVIDIA H100.
What is recorded
Each trace corresponds to one capture trial where:
A fixed expert id is selected (expert_00 ... expert_31).
A random hidden-state tensor is generated once per trial.
The selected expert computation is executed… See the full description on the dataset page: https://huggingface.co/datasets/masterpieceexternal/gpt-oss-20b-moe-expert-power-traces-320k.cs2_data_hf_expertfish_datasets_real_fowlers_expertsys_twodim_fourier_v2opencode_seed2.1_expert_skill_round_00opencode_seed2.1_expert_skill_round_00_20260712MoE_expert_selection_trace
📖 Introduction
This repository serves as a supplement to our paper "Patterns behind Chaos: Forecasting Data Movement for Efficient Large-Scale MoE LLM Inference".
It contains expert selection profiling traces of four top-tier MoE LLMs ranging from 235B to 1T (DeepSeek-R1, Kimi-K2-Thinking, Llama4-Marverick, and Qwen3-235B) across multiple benchmarks. For each query or request, we log the activated expert ID of every model layer of every generated token.
We provide analyses and… See the full description on the dataset page: https://huggingface.co/datasets/core12345/MoE_expert_selection_trace.glaucoma-expert-cot-final
Glaucoma Expert Chain-of-Thought
Ophthalmologist six-step reasoning reports for fundus photographs, each paired with
a binary glaucoma label. 1,074 cases from LAG and Papila.
Files
file
rows
glaucoma / not
train.jsonl
823
304 / 519
val.jsonl
92
46 / 46
test.jsonl
159
79 / 80
images/
1,074
<source>_<id>.jpg
Record schema
{
"id": "1689",
"source": "LAG",
"image": "LAG_1689.jpg",
"split": "train",
"final_diagnosis_GT":… See the full description on the dataset page: https://huggingface.co/datasets/yuzhench/glaucoma-expert-cot-final.opencode_seed2.1_expert_without_reproduce_round_00maniskill-dreamer4-expertqwen3.8-flash-next-expert-traces
Qwen3.8-Flash-Next expert routing traces
Token-level routing traces of a deployed MoE model: for every token and every one of the
48 MoE layers, which experts the router chose, the top-32 router logits behind that choice,
and the exact hidden state the router read — plus, in v3, the state at many layers per token,
the post-final-norm state the LM head consumes, and the LM head's top-8 next-token candidates.
The corpus exists to answer one question: how well can the next tokens'… See the full description on the dataset page: https://huggingface.co/datasets/aswinkumar99/qwen3.8-flash-next-expert-traces.kuka-centrifuge-expert-review_20260917_174048This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 30,
"features": {
"action": {
"dtype": "float32",
"names": [
"ee_x.pos",
"ee_y.pos",
"ee_z.pos",
"ee_wx.pos",
"ee_wy.pos",
"ee_wz.pos",
"gripper.pos"
],
"shape": [
7… See the full description on the dataset page: https://huggingface.co/datasets/ar0s/kuka-centrifuge-expert-review_20260917_174048.fish_datasets_real_electrodyn_expertsys_twodim_fourier_v2aihub_corpus_expertise
Dataset Card for "corpus_professional_field"
전문분야 말뭉치
ExpertHTR-Dataset
ExpertHTR Dataset
Gated page-level handwritten text recognition data for the
ExpertHTR project.
This is a rights-filtered replacement export: all HWDB/CASIA records and
images have been removed. The repository remains gated because the remaining
upstream sources have different access conditions. It is a companion data
release for ExpertHTR, not the exact training snapshot for the published
seven-source checkpoint.
Included data
Split
Records
Purpose… See the full description on the dataset page: https://huggingface.co/datasets/DAIR-Group/ExpertHTR-Dataset.gpt-oss-20b-moe-expert-power-traces-320k-ds16k
GPT-OSS-20B MoE Expert Power Traces (Downsampled to 16k)
Downsampled variant of the 320k expert-trace capture set.
Source
Raw source dataset (same captures):
32 experts (expert_00..expert_31)
10,000 traces per expert
320,000 total traces
raw trace length ~195k samples per trace
Downsampling method
Each raw trace was resampled to exactly 16384 samples using linear interpolation (np.interp) matching the trainer resampling step.
No baseline normalization and no… See the full description on the dataset page: https://huggingface.co/datasets/masterpieceexternal/gpt-oss-20b-moe-expert-power-traces-320k-ds16k.fathom-expert-dataExpertLongBench
🎓 ExpertLongBench: Expert-Level Benchmark for Long-Form Generation with Structured Checklists
📊 The leaderboard for ExpertLongBench is hosted here: 🔗 https://huggingface.co/spaces/launch/ExpertLongBench
This is the public portion of the ExpertLongBench dataset, introduced in the paper:
ExpertLongBench: Benchmarking Language Models on Expert-Level Long-Form Generation Tasks with Structured ChecklistsJie Ruan, Inderjeet Jayakumar Nair, Shuyang Cao, Amy Liu, Sheza Munir, Micah… See the full description on the dataset page: https://huggingface.co/datasets/launch/ExpertLongBench.Expert-Sudoku-100kexpertqa
Dataset Card for ExpertQA
Dataset Summary
We provide here the data accompanying the paper: ExpertQA: Expert-Curated Questions and Attributed Answers. The ExpertQA dataset contains 2177 examples from 32 different fields.
Supported Tasks
The main data contains 2177 examples that can be used to evaluate new methods for estimating factuality and attribution, while the lfqa_domain and lfqa_rand data can be used to evaluate long-form question answering systems.… See the full description on the dataset page: https://huggingface.co/datasets/cmalaviya/expertqa.calvin_expertsglaucoma-expert-cot-raw-1077
Glaucoma Expert Chain-of-Thought
Ophthalmologist six-step reasoning reports for fundus photographs, each paired with
a binary glaucoma label. 1,074 cases from LAG and Papila.
Files
file
rows
split
expert_cot_trainval.jsonl
915
train (823) + val (92)
expert_cot_test.jsonl
159
test
images/
1,074
<source>_<id>.jpg
Record schema
{
"id": "1689",
"source": "LAG",
"image": "LAG_1689.jpg",
"split": "train"… See the full description on the dataset page: https://huggingface.co/datasets/yuzhench/glaucoma-expert-cot-raw-1077.How-Resilient-are-Imitation-Learning-Methods-to-Sub-Optimal-Experts
How Resilient are Imitation Learning Methods to Sub-Optimal Experts?
Related Work
Trajectories used in How Resilient are Imitation Learning Methods to Sub-Optimal Experts?
The code that uses this data is on GitHub: https://github.com/NathanGavenski/How-resilient-IL-methods-are
Structure
These trajectories are formed by using Stable Baselines.
Each file is a dictionary of a set of trajectories with the following keys:
actions: the action in the given timestamp… See the full description on the dataset page: https://huggingface.co/datasets/NathanGavenski/How-Resilient-are-Imitation-Learning-Methods-to-Sub-Optimal-Experts.rl_expert_franka_dataset_for_pi0_1M_filtered_for_pickupsaime2024_Qwen3-30B-A3B_moe_patternsBUSTER
Dataset Card for BUSTER
BUSiness Transaction Entity Recognition dataset.
BUSTER is an Entity Recognition (ER) benchmark for entities related to business transactions. It consists of a gold corpus of
3779 manually annotated documents on financial transactions that were randomly divided into 5 folds,
plus an additional silver corpus of 6196 automatically annotated documents that were created by the model-optimized RoBERTa system.
Data Splits Statistics… See the full description on the dataset page: https://huggingface.co/datasets/expertai/BUSTER.math-500_Qwen3-30B-A3B_moe_patternsswe-oec-claude-expertgpqa_diamond_Qwen3-30B-A3B_moe_patterns
