CoolFace
20 results

BF16

brandonmusic /GLM-5.3-Flash-BF16-Teacher-Logits GLM-5.3-Flash BF16 teacher logits This dataset contains full-vocabulary float32 teacher logits from the immutable zai-org/GLM-5.3-Flash-BF16 revision a6c167b62691b2bac901344b65cb651a70f53e43. It keeps the sealed final KLD panel qualification-only and publishes the separate non-final calibration panel under role-specific paths. Qualification-only final windows: 25 Qualification-only final prediction positions: 51175 Vocabulary size: 154880 Teacher receipt:… See the full description on the dataset page: https://huggingface.co/datasets/brandonmusic/GLM-5.3-Flash-BF16-Teacher-Logits.text-generation4 likes5k downloads26d agoHugging Faceopen-llm-leaderboard-old /details_one-man-army__UNA-34Beagles-32K-bf16-v1 Dataset Card for Evaluation run of one-man-army/UNA-34Beagles-32K-bf16-v1 Dataset automatically created during the evaluation run of model one-man-army/UNA-34Beagles-32K-bf16-v1 on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_one-man-army__UNA-34Beagles-32K-bf16-v1.0 likes1.8k downloads3y agoHugging Faceapple /DataCompDR-12M-bf16 Dataset Card for DataCompDR-12M-BFloat16 This dataset contains synthetic captions, embeddings, and metadata for DataCompDR-12M. The metadata has been generated using pretrained image-text models on a 12M subset of DataComp-1B. For details on how to use the metadata, please visit our github repository. The dataset with the original captions is now available at mlfoundations/DataComp-12M. The UIDs per shards match between mlfoundations/DataComp-12M and apple/DataCompDR-12M-bf16.… See the full description on the dataset page: https://huggingface.co/datasets/apple/DataCompDR-12M-bf16.texttext-to-image10M<n<100M5 likes1.2k downloads5mo agoHugging Facebrandonmusic /GLM-5.3-BF16-full-logits0 likes1.1k downloads24d agoHugging Faceopen-llm-leaderboard-old /details_OpenBuddyEA__openbuddy-llama-30b-v7.1-bf16 Dataset Card for Evaluation run of OpenBuddyEA/openbuddy-llama-30b-v7.1-bf16 Dataset Summary Dataset automatically created during the evaluation run of model OpenBuddyEA/openbuddy-llama-30b-v7.1-bf16 on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_OpenBuddyEA__openbuddy-llama-30b-v7.1-bf16.0 likes366 downloads3y agoHugging FaceJWei05 /gemma4-e4b-rl100-hf-bf16-sdpa-topk128-overlay Gemma 4 E4B RL100 top-k-128 target overlay Precomputed off-policy distillation targets for the E4B-RL-step-100 to E2B experiment. Source traces: JWei05/gemma4-e4b-rl100-topk128-traces at revision 2b6e49a0a456ee9d67b16a1dc61785562bee90c9 Direction: Gemma 4 E4B RL step 100 teacher to Gemma 4 E2B base student Target engine: Hugging Face BF16 SDPA full forward Width: top-k 128 Stored target token IDs: int32 Stored target log-probabilities: float16 Causal alignment: response token… See the full description on the dataset page: https://huggingface.co/datasets/JWei05/gemma4-e4b-rl100-hf-bf16-sdpa-topk128-overlay.tabular10K<n<100K0 likes310 downloads2mo agoHugging Face