models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
Llama-3.2-3B_4bits_128group_sizetesting-llama_hidden_size_64bd3lm-owt-block_size4GLM-4.6-stackexchange-overflow-sandboxes-32eps-65k-reasoning_global-batch-size_32_Qwen3-32Bbd3lm-owt-block_size1024-pretrainbd3lm-owt-block_size16GLM-4.6-stackexchange-overflow-sandboxes-32eps-65k-reasoning_global-batch-size_64_Qwen3-32Bbd3lm-owt-block_size8Goedel-Code-Prover-8B-mlx-group_size64-quant_predicate_mixed_4_6multiplt-rkt-size-ablation-modelsEXAONE-Deep-2.4B-VPpos1_CoT0_Cri1_Hint0_Size500_2026-07-09BFS-Prover-V2-7B-mlx-group_size64-mixed_4_6Llama-8B-KNUT-ref-voice_size500_cot0_cri1_hint1opt-babylm2-rewritten-clean-spacy-earlystop_hierarchical_211_size-color_adj2-bpe_seed-211_1e-3Artigenz-Coder-DS-6.7B_dataset_size_52_epochs_10_2024-06-11_06-26-45_3520976babylama-hidden_sizes768unsloth-Qwen2.5-3B-Instruct-1_r-8_lr-0.0005_ms-25_gas-1_batch-size-8opt-babylm2-rewritten-clean-spacy-earlystop_hierarchical_211_size-color_adj2-bpe_seed-1024_1e-3codeparrot-ds-batch-size-2-gr-4test_model_eng_vocab_size_625deepseek-coder-6.7b-instruct_Fi__size_52_epochs_10_2024-06-21_03-01-23_3556388Llama-8B-KNUT-ref-voice_size500_cot1_cri1_hint1_retrainpopulation_size_extraction_bloomz3b_finetunecodellama-7b-hf-neuron-seqlen-2048-batch-size-2OpenCodeInterpreter-DS-6.7B_Fi__size_52_epochs_10_2024-06-21_02-53-25_3556384CodeLlama-7b-Instruct-hf_Fi__size_52_epochs_10_2024-06-21_04-40-42_3556390opt-babylm2-rewritten-clean-spacy-earlystop_hierarchical_211_size-origin_adj1-bpe_seed-42_1e-3opt-babylm2-rewritten-clean-spacy-earlystop_hierarchical_211_size-color_adj1-bpe_seed-211_1e-3opt-babylm2-rewritten-clean-spacy-earlystop_hierarchical_211_size-color_both-bpe_seed-211_1e-3unsloth-Qwen2.5-3B-Instruct-4_r-8_lr-0.0005_ms-25_gas-4_batch-size-8
