CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01serialexperimentsleon /fish_datasets_real_electrodyn_expertsys_twodim_fourier0 likes42k downloads1y agoHugging Face02mzxuan /robopro_expert0 likes37k downloads2mo agoHugging Face03masterpieceexternal /gpt-oss-20b-moe-expert-power-traces-320k GPT-OSS-20B MoE Expert Power Traces (320k, ChipWhisperer) This dataset contains analog power traces captured with a ChipWhisperer Husky while running forced single-expert MoE computations derived from openai/gpt-oss-20b on an NVIDIA H100. What is recorded Each trace corresponds to one capture trial where: A fixed expert id is selected (expert_00 ... expert_31). A random hidden-state tensor is generated once per trial. The selected expert computation is executed… See the full description on the dataset page: https://huggingface.co/datasets/masterpieceexternal/gpt-oss-20b-moe-expert-power-traces-320k.audio-classification100K<n<1M0 likes5.8k downloads4mo agoHugging Face04tuan124816 /cs2_data_hf_expert10K<n<100K0 likes5k downloads2y agoHugging Face05serialexperimentsleon /fish_datasets_real_fowlers_expertsys_twodim_fourier_v20 likes3.5k downloads1y agoHugging Face06JianhuiWei /opencode_seed2.1_expert_skill_round_000 likes3.1k downloads2mo agoHugging Face07JianhuiWei /opencode_seed2.1_expert_skill_round_00_20260712image0 likes2k downloads2mo agoHugging Face08core12345 /MoE_expert_selection_tracegated 📖 Introduction This repository serves as a supplement to our paper "Patterns behind Chaos: Forecasting Data Movement for Efficient Large-Scale MoE LLM Inference". It contains expert selection profiling traces of four top-tier MoE LLMs ranging from 235B to 1T (DeepSeek-R1, Kimi-K2-Thinking, Llama4-Marverick, and Qwen3-235B) across multiple benchmarks. For each query or request, we log the activated expert ID of every model layer of every generated token. We provide analyses and… See the full description on the dataset page: https://huggingface.co/datasets/core12345/MoE_expert_selection_trace.26 likes1.6k downloads4mo agoHugging Face09yuzhench /glaucoma-expert-cot-final Glaucoma Expert Chain-of-Thought Ophthalmologist six-step reasoning reports for fundus photographs, each paired with a binary glaucoma label. 1,074 cases from LAG and Papila. Files file rows glaucoma / not train.jsonl 823 304 / 519 val.jsonl 92 46 / 46 test.jsonl 159 79 / 80 images/ 1,074 <source>_<id>.jpg Record schema { "id": "1689", "source": "LAG", "image": "LAG_1689.jpg", "split": "train", "final_diagnosis_GT":… See the full description on the dataset page: https://huggingface.co/datasets/yuzhench/glaucoma-expert-cot-final.imageimage-classificationn<1K0 likes1.1k downloads2mo agoHugging Face10JianhuiWei /opencode_seed2.1_expert_without_reproduce_round_00text0 likes779 downloads2mo agoHugging Face11MinghaoFu /maniskill-dreamer4-expert0 likes656 downloads4mo agoHugging Face12aswinkumar99 /qwen3.8-flash-next-expert-traces Qwen3.8-Flash-Next expert routing traces Token-level routing traces of a deployed MoE model: for every token and every one of the 48 MoE layers, which experts the router chose, the top-32 router logits behind that choice, and the exact hidden state the router read — plus, in v3, the state at many layers per token, the post-final-norm state the LM head consumes, and the LM head's top-8 next-token candidates. The corpus exists to answer one question: how well can the next tokens'… See the full description on the dataset page: https://huggingface.co/datasets/aswinkumar99/qwen3.8-flash-next-expert-traces.text-generation2 likes593 downloads12d agoHugging Face13ar0s /kuka-centrifuge-expert-review_20260917_174048This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "fps": 30, "features": { "action": { "dtype": "float32", "names": [ "ee_x.pos", "ee_y.pos", "ee_z.pos", "ee_wx.pos", "ee_wy.pos", "ee_wz.pos", "gripper.pos" ], "shape": [ 7… See the full description on the dataset page: https://huggingface.co/datasets/ar0s/kuka-centrifuge-expert-review_20260917_174048.tabularrobotics100K<n<1M0 likes424 downloads4d agoHugging Face14serialexperimentsleon /fish_datasets_real_electrodyn_expertsys_twodim_fourier_v20 likes383 downloads1y agoHugging Face15wisenut-nlp-team /aihub_corpus_expertise Dataset Card for "corpus_professional_field" 전문분야 말뭉치 textother10M<n<100M1 likes362 downloads3y agoHugging Face16DAIR-Group /ExpertHTR-Datasetgated ExpertHTR Dataset Gated page-level handwritten text recognition data for the ExpertHTR project. This is a rights-filtered replacement export: all HWDB/CASIA records and images have been removed. The repository remains gated because the remaining upstream sources have different access conditions. It is a companion data release for ExpertHTR, not the exact training snapshot for the published seven-source checkpoint. Included data Split Records Purpose… See the full description on the dataset page: https://huggingface.co/datasets/DAIR-Group/ExpertHTR-Dataset.imageimage-to-text10K<n<100K1 likes362 downloads6d agoHugging Face17masterpieceexternal /gpt-oss-20b-moe-expert-power-traces-320k-ds16k GPT-OSS-20B MoE Expert Power Traces (Downsampled to 16k) Downsampled variant of the 320k expert-trace capture set. Source Raw source dataset (same captures): 32 experts (expert_00..expert_31) 10,000 traces per expert 320,000 total traces raw trace length ~195k samples per trace Downsampling method Each raw trace was resampled to exactly 16384 samples using linear interpolation (np.interp) matching the trainer resampling step. No baseline normalization and no… See the full description on the dataset page: https://huggingface.co/datasets/masterpieceexternal/gpt-oss-20b-moe-expert-power-traces-320k-ds16k.audio-classification100K<n<1M0 likes356 downloads7mo agoHugging Face18umer07 /fathom-expert-data0 likes356 downloads6mo agoHugging Face19launch /ExpertLongBench 🎓 ExpertLongBench: Expert-Level Benchmark for Long-Form Generation with Structured Checklists 📊 The leaderboard for ExpertLongBench is hosted here: 🔗 https://huggingface.co/spaces/launch/ExpertLongBench This is the public portion of the ExpertLongBench dataset, introduced in the paper: ExpertLongBench: Benchmarking Language Models on Expert-Level Long-Form Generation Tasks with Structured ChecklistsJie Ruan, Inderjeet Jayakumar Nair, Shuyang Cao, Amy Liu, Sheza Munir, Micah… See the full description on the dataset page: https://huggingface.co/datasets/launch/ExpertLongBench.text-generation10 likes354 downloads1y agoHugging Face20Alignment-Lab-AI /Expert-Sudoku-100ktabular100K<n<1M0 likes347 downloads2y agoHugging Face21cmalaviya /expertqa Dataset Card for ExpertQA Dataset Summary We provide here the data accompanying the paper: ExpertQA: Expert-Curated Questions and Attributed Answers. The ExpertQA dataset contains 2177 examples from 32 different fields. Supported Tasks The main data contains 2177 examples that can be used to evaluate new methods for estimating factuality and attribution, while the lfqa_domain and lfqa_rand data can be used to evaluate long-form question answering systems.… See the full description on the dataset page: https://huggingface.co/datasets/cmalaviya/expertqa.textquestion-answering1K<n<10K11 likes328 downloads3y agoHugging Face22mattewg /calvin_expertstextn<1K1 likes327 downloads6mo agoHugging Face23yuzhench /glaucoma-expert-cot-raw-1077 Glaucoma Expert Chain-of-Thought Ophthalmologist six-step reasoning reports for fundus photographs, each paired with a binary glaucoma label. 1,074 cases from LAG and Papila. Files file rows split expert_cot_trainval.jsonl 915 train (823) + val (92) expert_cot_test.jsonl 159 test images/ 1,074 <source>_<id>.jpg Record schema { "id": "1689", "source": "LAG", "image": "LAG_1689.jpg", "split": "train"… See the full description on the dataset page: https://huggingface.co/datasets/yuzhench/glaucoma-expert-cot-raw-1077.imageimage-classificationn<1K0 likes313 downloads2mo agoHugging Face24NathanGavenski /How-Resilient-are-Imitation-Learning-Methods-to-Sub-Optimal-Experts How Resilient are Imitation Learning Methods to Sub-Optimal Experts? Related Work Trajectories used in How Resilient are Imitation Learning Methods to Sub-Optimal Experts? The code that uses this data is on GitHub: https://github.com/NathanGavenski/How-resilient-IL-methods-are Structure These trajectories are formed by using Stable Baselines. Each file is a dictionary of a set of trajectories with the following keys: actions: the action in the given timestamp… See the full description on the dataset page: https://huggingface.co/datasets/NathanGavenski/How-Resilient-are-Imitation-Learning-Methods-to-Sub-Optimal-Experts.other100B<n<1T0 likes312 downloads4y agoHugging Face25raulsteleac /rl_expert_franka_dataset_for_pi0_1M_filtered_for_pickupsimage1M<n<10M0 likes312 downloads1y agoHugging Face26ExpertFlowPredictor /aime2024_Qwen3-30B-A3B_moe_patternstext10K<n<100K0 likes303 downloads11mo agoHugging Face27expertai /BUSTER Dataset Card for BUSTER BUSiness Transaction Entity Recognition dataset. BUSTER is an Entity Recognition (ER) benchmark for entities related to business transactions. It consists of a gold corpus of 3779 manually annotated documents on financial transactions that were randomly divided into 5 folds, plus an additional silver corpus of 6196 automatically annotated documents that were created by the model-optimized RoBERTa system. Data Splits Statistics… See the full description on the dataset page: https://huggingface.co/datasets/expertai/BUSTER.texttoken-classification1K<n<10K5 likes288 downloads2y agoHugging Face28ExpertFlowPredictor /math-500_Qwen3-30B-A3B_moe_patternstext10K<n<100K0 likes285 downloads11mo agoHugging Face29ScaleAI /swe-oec-claude-experttext1K<n<10K1 likes285 downloads11mo agoHugging Face30ExpertFlowPredictor /gpqa_diamond_Qwen3-30B-A3B_moe_patternstext10K<n<100K0 likes242 downloads11mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.