models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
deep-ignorance-pretraining-stage-unfiltereddeep-ignorance-pretraining-stage-strong-filtertransformer-pretraining-from-scratchQwen2-1.5B-confluence-pretrainingdeep-ignorance-pretraining-stage-weak-filterpretraining_testpoker-pretrainingpretraining-priors-doormail4k-d32-treated-baseqwen3-micro-1kmistral-pretrainingpre-pretraining-10Mf-gpt2-small-42mistral-pretraining100k_fineweb_continued_pretraining_Qwen2.5-0.5B-Instruct_Unsloth_merged_16bitpretraining-priors-d26-basevicuna-7b-v1.3-tiny-stories-pretraining-2epochdeep_ignorance_pretraining_baseline_smalldeep_ignorance_pretraining_filtered_smallpretraining-priors-d26-sftpretraining-priors-d26-sft-numtoxpretraining-priors-d26-base-numtox83k_continued_pretraining_Qwen2.5-0.5B-Instruct_Unsloth_merged_16bitQwen2.5-Coder-1.5B-Instruct_MATH_training_response_Qwen_QwQ_32B_Preview_common_correct_leveltrellis-pretrainingQwen2.5-Coder-1.5B-Instruct_MATH_training_Qwen_QwQ_32B_Preview10k_continued_pretraining_Qwen2.5-0.5B-Instruct_Unsloth_merged_16bittest_continued_pretraining_Phi-3-mini-4k-instruct_Unsloth_merged_16bitpre_training_llama10k_continued_pretraining_Phi-3-mini-4k-instruct_Unsloth_merged_16bit40k_continued_pretraining_Qwen2.5-0.5B-Instruct_Unsloth_merged_16bitmistral-pretraining
