CoolFace
15 results

llada2

TIDE-dllm /distill_llada2_sft distill_llada2_sft — Pre-tokenized SFT mixture for LLaDA2-teacher distillation Pre-tokenized SFT corpus used to train every checkpoint in the Cross-Tokenizer (Pipeline A) of the TIDE framework — i.e. the distill-LLaDA2-* student checkpoints distilled from inclusionAI/LLaDA2.0-mini. The dataset ships as a datasets.DatasetDict (load_from_disk-ready) so distillation training never has to re-tokenize at job start (which would cause NCCL timeouts on multi-node runs).… See the full description on the dataset page: https://huggingface.co/datasets/TIDE-dllm/distill_llada2_sft.text-generation1M<n<10M0 likes156 downloads5mo agoHugging Faceweizhou03 /Llada2_EB_Decode_512_sequential_dataset0 likes77 downloads2mo agoHugging Facekarantonis /llada2-mini-uq-pickles LLaDA 2.1 mini — UQ eval pickles (ue_manager_seed1) Durable backup of the LLaDA 2.1 mini uncertainty-quantification sweep pickles produced with lm-polygraph (branch feat/llada2-cache, April–May 2026). Each .pkl is a torch.load-able ue_manager dump containing per-sample stats (greedy_texts, target_texts, ...), estimations, gen_metrics, and metrics. Layout setup_a/ no-train baselines, reduced config (~34 est, K=10 sampling dropped) setup_b/ with-train baselines… See the full description on the dataset page: https://huggingface.co/datasets/karantonis/llada2-mini-uq-pickles.0 likes63 downloads3mo agoHugging Facezimplex /dllm-attention-parents-llada2-finecode-pred-debias5-b64-v1 DART FineCode deterministic 1/8 sample — inclusionAI/LLaDA2.0-mini dependency sidecar This is the sealed attention dependency release generated with pred-debias5 corrected attention. It is compatible with block size 64 and pairs only with zimplex/dllm-finecode-dagcover-full-v1, whose manifest SHA-256 is 45e250f9f4a5302ef1bb04c16577eb8b6d85996368fe1dc359b86465cdb03176. The sidecar manifest SHA-256 is f4dbab97f495e23682f26ac085b19a90e9e6a15cf664b789d9efb0f23566cd3e. The source… See the full description on the dataset page: https://huggingface.co/datasets/zimplex/dllm-attention-parents-llada2-finecode-pred-debias5-b64-v1.0 likes53 downloads7d agoHugging Facezimplex /dllm-effect-parents-llada2-finemath-half-d1marginal-w128-tau2-b64-v3 DART FineMath-4+ half — inclusionAI/LLaDA2.0-mini dependency sidecar This is the sealed counterfactual dependency release generated with corrected d1-marginal, window 128, tau 2. It is compatible with block size 64 and pairs only with zimplex/dllm-finemath-dagcover-half-v1, whose manifest SHA-256 is e20f3ecdc7417e8f8141610d4e1c55eb2973d269d03117d47c4e27b3d1182a68. The sidecar manifest SHA-256 is fbf0371aeb9de796c5b8d22730e91c4254faef1354151ac797dc20671a2a848a. The source release… See the full description on the dataset page: https://huggingface.co/datasets/zimplex/dllm-effect-parents-llada2-finemath-half-d1marginal-w128-tau2-b64-v3.0 likes50 downloads7d agoHugging Facezimplex /dllm-effect-parents-llada2-finecode-d1marginal-w128-tau2-b64-v3 DART FineCode deterministic 1/8 sample — inclusionAI/LLaDA2.0-mini dependency sidecar This is the sealed counterfactual dependency release generated with corrected d1-marginal, window 128, tau 2. It is compatible with block size 64 and pairs only with zimplex/dllm-finecode-dagcover-full-v1, whose manifest SHA-256 is 45e250f9f4a5302ef1bb04c16577eb8b6d85996368fe1dc359b86465cdb03176. The sidecar manifest SHA-256 is 391fa0f852451457ba2812a8633af7799f18193e5e0d2a489432ccc7c3d67c43.… See the full description on the dataset page: https://huggingface.co/datasets/zimplex/dllm-effect-parents-llada2-finecode-d1marginal-w128-tau2-b64-v3.0 likes46 downloads7d agoHugging Face