datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Distillation_RAGdistillation-llm-rawdistillationDescription: Snapshot measurements on 27 variables from a distillation column; measured over 2.5 years.
Data source: From an industrial source; variable names have been coded. e.g. Temp1 is a temperature, but we cannot disclose where it is measured on the column.
Temperatures are in Fahrenheit
Pressures are measured in bars
FlowC1 in units of MSCFD
FlowC3 and FlowC4 are in units of MBPD
Temp11 = Temp3 - Temp9 = the temperature increase of the stream leaving the column and returning back, after… See the full description on the dataset page: https://huggingface.co/datasets/talaviyabhavik/distillation.Grok-Code-Fast-1-Distillation-Done-By-GPT5.4
GPT 5.4 Code Distillation
This dataset contains 500 randomly sampled prompt, reasoning, and output triples derived from the source dataset TeichAI/grok-code-fast-1-1000x.
Columns
Prompt: the user message extracted from the source conversation.
Reasoning: the assistant's <think> content when present.
Output: the assistant response after the <think> block.
Source Attribution
Prompt source and original conversation data come from TeichAI/grok-code-fast-1-1000x.… See the full description on the dataset page: https://huggingface.co/datasets/SLoonker/Grok-Code-Fast-1-Distillation-Done-By-GPT5.4.
