CoolFace
20 results

moe

moenz /hf-9router2 likes9.7k downloads1d agoHugging Facemasterpieceexternal /gpt-oss-20b-moe-expert-power-traces-320k GPT-OSS-20B MoE Expert Power Traces (320k, ChipWhisperer) This dataset contains analog power traces captured with a ChipWhisperer Husky while running forced single-expert MoE computations derived from openai/gpt-oss-20b on an NVIDIA H100. What is recorded Each trace corresponds to one capture trial where: A fixed expert id is selected (expert_00 ... expert_31). A random hidden-state tensor is generated once per trial. The selected expert computation is executed… See the full description on the dataset page: https://huggingface.co/datasets/masterpieceexternal/gpt-oss-20b-moe-expert-power-traces-320k.audio-classification100K<n<1M0 likes5.8k downloads4mo agoHugging FaceMoeNew /OmniFake OmniFake OmniFake is a large-scale, well-categorized synthetic image dataset introduced in Few-Shot Synthetic Image Attribution: Identifying Unseen Generators with Limited Samples. It contains 1.17 million AI-generated images from 45 distinct generators, paired with 1.17 million real images, designed for research on AI-generated image (AIGI) detection and source attribution. For usage instructions and experimental protocols, please refer to the OmniDFA GitHub repository.… See the full description on the dataset page: https://huggingface.co/datasets/MoeNew/OmniFake.imageimage-classification1M<n<10M0 likes3k downloads3mo agoHugging FacesunLry /moe-demo-clean RoboTwin MOE Demo Clean Raw RoboTwin demonstration data copied from bos:/lab-test/moe-demo-clean/. The dataset is organized by task directories such as place_can_basket-demo_clean-200/. Each task directory contains episode subdirectories with an HDF5 trajectory file and an instructions.json file. text1K<n<10K0 likes2.5k downloads3mo agoHugging Facemarin-community /grug-moe-mix-swarm Grug-MoE Data-Mix Experiments The default config contains the original 840-run Fisher-DSP swarm. The harrier_18t75_d768 config contains the Harrier experiments described below. Fisher-DSP swarm (default) 840 MoE pretraining runs from the Grug-MoE Fisher-DSP data-mixing swarm (d512, TPU / us-central2). Each run trains on a distinct data mixture over 168 datakit buckets; the swarm is used to regress mixture weights → eval loss and predict an optimized pretraining… See the full description on the dataset page: https://huggingface.co/datasets/marin-community/grug-moe-mix-swarm.tabulartext-generation1K<n<10K1 likes1.9k downloads10d agoHugging FaceMoE-UNC /story_clozetext1K<n<10K1 likes1.4k downloads3y agoHugging Face