datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
DataCompDR-12M-bf16
Dataset Card for DataCompDR-12M-BFloat16
This dataset contains synthetic captions, embeddings, and metadata for DataCompDR-12M.
The metadata has been generated using pretrained image-text models on a 12M subset of DataComp-1B.
For details on how to use the metadata, please visit our github repository.
The dataset with the original captions is now available at mlfoundations/DataComp-12M.
The UIDs per shards match between mlfoundations/DataComp-12M and apple/DataCompDR-12M-bf16.… See the full description on the dataset page: https://huggingface.co/datasets/apple/DataCompDR-12M-bf16.gemma4-e4b-rl100-hf-bf16-sdpa-topk128-overlay
Gemma 4 E4B RL100 top-k-128 target overlay
Precomputed off-policy distillation targets for the E4B-RL-step-100 to E2B experiment.
Source traces: JWei05/gemma4-e4b-rl100-topk128-traces at revision 2b6e49a0a456ee9d67b16a1dc61785562bee90c9
Direction: Gemma 4 E4B RL step 100 teacher to Gemma 4 E2B base student
Target engine: Hugging Face BF16 SDPA full forward
Width: top-k 128
Stored target token IDs: int32
Stored target log-probabilities: float16
Causal alignment: response token… See the full description on the dataset page: https://huggingface.co/datasets/JWei05/gemma4-e4b-rl100-hf-bf16-sdpa-topk128-overlay.details_Kquant03__CognitiveFusion2-4x7B-BF16
Dataset Card for Evaluation run of Kquant03/CognitiveFusion2-4x7B-BF16
Dataset automatically created during the evaluation run of model Kquant03/CognitiveFusion2-4x7B-BF16.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_Kquant03__CognitiveFusion2-4x7B-BF16.lm-eval-results-Kquant03-Nanashi-2x7B-bf16-private
Dataset Card for Evaluation run of Kquant03/Nanashi-2x7B-bf16
Dataset automatically created during the evaluation run of model Kquant03/Nanashi-2x7B-bf16
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-Kquant03-Nanashi-2x7B-bf16-private.lm-eval-results-Kquant03-Cognito-2x7B-bf16-private
Dataset Card for Evaluation run of Kquant03/Cognito-2x7B-bf16
Dataset automatically created during the evaluation run of model Kquant03/Cognito-2x7B-bf16
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-Kquant03-Cognito-2x7B-bf16-private.BF16kEval_FinEval_16k_fulleval__3args_ours-eval_rllm-eval-results-CultriX-NeuralTrix-bf16-private
Dataset Card for Evaluation run of CultriX/NeuralTrix-bf16
Dataset automatically created during the evaluation run of model CultriX/NeuralTrix-bf16
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.
An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-CultriX-NeuralTrix-bf16-private.Qwen3.8-DSpark-PerfectBlend-5M-Paired-BF16
Qwen3.8 DSpark PerfectBlend 5M paired BF16 features
Private, checksum-closed paired feature corpus for the Qwen3.8 Flash / 27B
DSpark transplant project. Repository: MJPansa/Qwen3.8-DSpark-PerfectBlend-5M-Paired-BF16.
The Hugging Face DatasetDict rows are a compact index. Each row points into
three immutable SafeTensor files in tensors/shard-NNNNN/ using exact token
and anchor offsets. This keeps the ~180 GB dense BF16 corpus resumable and
memory-mappable instead of duplicating… See the full description on the dataset page: https://huggingface.co/datasets/MJPansa/Qwen3.8-DSpark-PerfectBlend-5M-Paired-BF16.BF16kEval_FinEval_16k_fulleval__3args_r1-eval_0BF16kEval_FinEval_16k_fulleval__3args_rlonly-eval_rleval-NVIDIA-Nemotron-3-Nano-30B-A3B-BF16_16concurrency_eval_ctx131k_terminal-bench-2.0lm-eval-results-CultriX-NeuralTrixlaser-bf16-private
Dataset Card for Evaluation run of CultriX/NeuralTrixlaser-bf16
Dataset automatically created during the evaluation run of model CultriX/NeuralTrixlaser-bf16
The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-CultriX-NeuralTrixlaser-bf16-private.terminal_bench_2_NVIDIA_Nemotron_3_Nano_30B_A3B_BF16_20260415_174206GLM-5.2-BF16-KLD-Reference-Logits-20260618
GLM-5.2 BF16 KLD Reference Logits
Reference logits for local GLM-5.2 KLD checks.
Contents:
prefill/logits_0.safetensors: BF16 prefill prompt logits generated from
zai-org/GLM-5.2 with context length 2048, stride 512, one window.
decode/decode_teacher_bf16_ref_ctx2048_t17_20260618.safetensors: BF16
teacher-forced decode logits for prompt length 2048 and 17 decode tokens.
decode/decode_teacher_bf16_ref_ctx2048_t17_20260618.safetensors.json:
metadata for the decode reference.… See the full description on the dataset page: https://huggingface.co/datasets/festr2/GLM-5.2-BF16-KLD-Reference-Logits-20260618.eval-NVIDIA-Nemotron-3-Nano-30B-A3B-BF16_16concurrency_openhands_eval_c_terminal67fe5eedGLM-5.2-BF16-KLD-Reference-Logits-20260708
GLM-5.2 BF16 KLD Reference Logits 20260708
This dataset contains the current GLM-5.2 BF16 reference prompt logits used for
the July 2026 vLLM/Blackwell KLD checks.
The reference cache is intended for candidate-side KLD comparisons without
rerunning the expensive BF16 reference pass.
Files
reference-logits/logits_0.safetensors
reference-logits/manifest.json
generation-log/config.env
generation-log/scoremode_kld.log
Reference Generation
Field… See the full description on the dataset page: https://huggingface.co/datasets/festr2/GLM-5.2-BF16-KLD-Reference-Logits-20260708.eurospeech-latents-bf16terminal_bench_2_NVIDIA_Nemotron_3_Nano_30B_A3B_BF16_20260414_211305gaia_127_NVIDIA_Nemotron_3_Nano_30B_A3B_BF16_20260430_193800eval-NVIDIA-Nemotron-3-Nano-30B-A3B-BF16_16concurrency_eval_ctx131k_OpenThoughts-TB-devterminal_bench_2_rl__nemotron_bash_bf16_terminus_2_32b_20260407_232107terminal_bench_2_NVIDIA_Nemotron_3_Nano_30B_A3B_BF16_20260425_155129summarize_from_feedback_tldr3_unlabelled_vllm_dpo_costa_2.8b_bf16.yml_6e799_newdev_set_v2_NVIDIA_Nemotron_3_Nano_30B_A3B_BF16_20260425_155108-traceseval-NVIDIA-Nemotron-3-Nano-30B-A3B-BF16_16concurrency_eval_ctx131k_swebench-vera9b71b18dev_set_v2_NVIDIA_Nemotron_3_Nano_30B_A3B_BF16_20260414_211234Qwen3.6-27B-AWQ-BF16-INT4-SuperGPQA-benchmarkBenchmark of cyankiwi/Qwen3.6-27B-AWQ-BF16-INT4 against m-a-p/SuperGPQA dataset.
Accuracy: 69.2% with Python tool.
Metric
Value
Correct
692
Incorrect
295
Errors
13
Total samples
1000
Python tool calls
1508
Total completion tokens
3,806,045
Raw stats:
{
"accuracy": 0.692,
"correct": 692,
"incorrect": 295,
"error": 13,
"total": 1000,
"python_tool_calls": 1508,
"completion_tokens": 3806045
}
gaia_127_NVIDIA_Nemotron_3_Nano_30B_A3B_BF16_20260425_071200eval-NVIDIA-Nemotron-3-Nano-30B-A3B-BF16_16concurrency_openhands_eval_c_OpenThoue429c793dev_set_v2_NVIDIA_Nemotron_3_Nano_30B_A3B_BF16_20260415_174135
