models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
Hugston_code-rl-Qwen3-4B-Instruct-2507-SFT-30bGRM-Gemma-2B-sftregLlama-3.1-Tulu-3-8B-SFT-RM-RB2Llama-3.1-Tulu-3-70B-SFT-RM-RB2GRM-llama3-8B-sftregpythia1b-sft-rm-tldrpromptFirstName-Genre-Classifier-30M-SFTSFT_anthropic_hh_EleutherAI1bGRM-Gemma2-2B-sftregGRM-llama3.2-3B-sftregrlhflow-llama-3-sft-8b-v2-segment-rm-700kKomdigiITS-3B-PAD-SFTPhishMe-R1-8B-SFTfinbert-full-sft-financial-sentiment_v3modernbert-ja-310m-sftxlm-mlm-plains-cree-en-silver-sft-hierarchicalmistral-7b-sft-beta__100000_1e-05_RewardModel_2GPUllama3.2-lora-sft-perfume-classificationRM_Mistral_sft_init_ultrafeedbck_lr_5e6summarization_sft_reward-model-deberta-v3-large-v2_RM-Gemma-2B_mask_partial_rm_random_lengthsft_reward_model_finalgalactica-6.7b-SFT-Rerank-GSM8kLlama-2-7b-chat-hf_mixed_sft_lexical_instruction_final_fullpythia-70m_tatsu-lab_alpaca_farm_sftsd1_policy_pythia-6.9b_gold_internlm2-7b_noise0.25_rmsd1SFT_GradProjectsummarization_sft_reward-model-deberta-v3-large-v2PhishMe-Qwen3-Base-8B-SFTLlama-3.2-1B-Instruct-SFT-EN_VNCodellama-7b-hf-SFT-Rerank-GSM8kMeta-Llama-3-8B-Instruct_mixed_sft_lexical_instruction_final_full
