models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
correctional-GPTaMLearningToPresent-RL-Qwen-2.5B-Coder-Instruct-GRPO-FinetunedCoody-learningBudyfinetuned_qwen_kanji_learningqwen_3_0_6b_sandbagging_linguistics_and_language_learning_sandbagging_2_epocheval3_phase2_TOYQwen3-4B-Thinking-2507-Heretic-CodeFeedback-OpenCodeInstruct-Learning-LoRAqwen3-14b-sft-5e-6-learning-rateqwen3-14b-sft-1e-5-learning-rate-h4codellama-7b-learning_rate2e-4eval3_lora_fixed_Chengmingfinance_stage1_adapterqwen3-14b-sft-1e-5-learning-ratesubliminal-learning-tiger-bothqwen_3_4b_sandbagging_linguistics_and_language_learning_sandbagging_2_epocheval3_yann_left_repair_adaptereval2_vqa_lora_parsefixgpt2-lora-ai-demofinance_sft_adapterfinance_dpo_adapterLearning_Model_medical_FineTuningqwen3-sft-learningrate1e3eval2_resolver_v2_besteval3_phase1_celebeval3_phase1_TOYeval3_phase2_celeb
