datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
grok-1-goz1-packs
Grok-1 GOZ1 ternary weight packs (research)
Packaged by: Grok Build: Grok 4.5 (high)
Research binary packs produced by rmems/grok-ozempic
from open xai-org/grok-1 ckpt-0 weights
(Apache-2.0). Format is GOZ1 (custom ternary/SNN container) — not
safetensors, not GGUF, not loadable by transformers.
Companion experiment reports & metrics:
rmems/grok-1-ternary-quant-experiments.
Structural inventory (no weights):
rmems/grok-1-dissect-artifacts.
What is included… See the full description on the dataset page: https://huggingface.co/datasets/rmems/grok-1-goz1-packs.grok-1-dissect-artifacts
Grok-1 structural parse artifacts (personal research)
Personal research tooling output — not a product release and not a hybrid-quantization
implementation. Structural parse of open Grok-1 checkpoint shards (tensor inventory, MoE
expert layout, routing-critical tensors, conversion policies, reports) produced by the
rmems/xai-dissect weight parser.
This is not a copy of the model weights, not a finetune, not an inference
runtime, and does not implement hybrid quantization.… See the full description on the dataset page: https://huggingface.co/datasets/rmems/grok-1-dissect-artifacts.grok-1-ternary-quant-experiments
Grok-1 SAAQ quantization / route-preservation experiments
Dataset author: Raul Montoya Cardenas (rmems)
SAAQ stands for Spiking Adaptive Activity Quantization, a term coined by
the dataset author.
Attribution: Grok Build: Grok 4.5 (high) packaged the original 2026-08-10
dataset. Codex: GPT-5.6-Sol (OpenAI) implemented, executed, validated, and published the
canonical issue #85 v4 evidence added on 2026-08-24.
Personal research measuring route preservation when packing open… See the full description on the dataset page: https://huggingface.co/datasets/rmems/grok-1-ternary-quant-experiments.tasklist-grok4-multilingual-50000x-unfiltered
TaskGen Dataset
Generated with taskgen by empero-org
Run Parameters
Parameter
Value
Model
grok-4-1-fast-non-reasoning
Temperature
0.75
Total Tasks
50000
Concurrency
8 workers
API Base
https://api.x.ai/v1
Generated
2026-04-07 09:04:57
Budget Cap
$15.0000
Multilingual
Yes (en, de, fr, es, nl, zh, ar, ru)
Language Distribution
Language
Code
Tasks
Arabic
ar
6111
Chinese
zh
6058
German
de
6057
Spanish
es
6020… See the full description on the dataset page: https://huggingface.co/datasets/empero-ai/tasklist-grok4-multilingual-50000x-unfiltered.tasklist-grok-multilingual-100000x-unfiltered
TaskGen Dataset
Generated with taskgen by empero-org
Run Parameters
Parameter
Value
Model
grok-4-1-fast-reasoning
Temperature
0.9
Total Tasks
83052
Concurrency
30 workers
API Base
https://api.x.ai/v1
Generated
2026-04-07 14:31:14
Budget Cap
$15.0000
Multilingual
Yes (en, de, fr, es, nl, zh, ar, ru)
Language Distribution
Language
Code
Tasks
Arabic
ar
10446
German
de
10397
Dutch
nl
10353
Spanish
es
10345… See the full description on the dataset page: https://huggingface.co/datasets/empero-ai/tasklist-grok-multilingual-100000x-unfiltered.grok-reflection-cot-ru
march228/grok-reflection-cot-ru
Russian synthetic dataset with question, internal thought text, and final answer.
What is inside
Rows: 4190
Split: train
Main fields:
question
thought_text
answer
thought1..thought5
model
task_type
reflection_count
Format
The dataset is stored as train.jsonl.
thought_text is the joined internal monologue with blank lines between thought blocks.thought1..thought5 preserve the original segmented form from the SQLite source.… See the full description on the dataset page: https://huggingface.co/datasets/march228/grok-reflection-cot-ru.spectral-smear-groktradehax-xai-grok-trading-visual-prompts
TradeHax xAI/Grok Trading Visual Prompts
Curated trading-scene image prompts and negative prompts tuned for xAI/Grok-inspired visual style.
Owner: Antired
Repo: tradehax-xai-grok-trading-visual-prompts
Synced by: scripts/sync-hf-assets.js
grok-grokking-ridge-repro-code
