models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
math-shepherd-mistral-7b-prm-calibrated-Llama-3.2-1B-InstructQwen2.5-Math-PRM-7B-calibrated-Llama-3.2-1B-Instructmath-shepherd-mistral-7b-prm-calibrated-Qwen2.5-Math-1.5B-InstructQwen2.5-Math-PRM-7B-calibrated-Qwen2.5-Math-1.5B-InstructMistral_nemo_calibrated_f1enhanced_full_oldinstruct_bestMistral_nemo_calibrated_f1enhanced_full_oldinstruct_final_v3Qwen2.5-Math-PRM-7B-calibrated-Qwen2.5-Math-7B-InstructReasonEval-7B-calibrated-DeepSeek-R1-Distill-Qwen-7BReasonEval-7B-calibrated-DeepSeek-R1-Distill-Llama-8BMistral_nemo_calibrated_f1enhanced_10shot_oldinstruct_bestv2tower-calibratedQwen2.5-Math-PRM-7B-calibrated-DeepSeek-R1-Distill-Qwen-7Bmath-shepherd-mistral-7b-prm-calibrated-Qwen2.5-Math-7B-InstructReasonEval-7B-calibrated-Qwen2.5-Math-1.5B-InstructQwen2.5-Math-PRM-7B-calibrated-Llama-3.1-8B-InstructQwen2.5-Math-PRM-7B-calibrated-DeepSeek-R1-Distill-Llama-8Bmath-shepherd-mistral-7b-prm-calibrated-DeepSeek-R1-Distill-Llama-8Bmath-shepherd-mistral-7b-prm-calibrated-DeepSeek-R1-Distill-Qwen-7BReasonEval-7B-calibrated-Qwen2.5-Math-7B-Instructcircuit-detective-qwen35-2b-phase2-calibrated-a10g-loracalibrated-research-qa-judgeMistral_7b_penalty_calibrated_finetuning_bestMistral_nemo_calibrated_f1enhanced_full_oldinstruct_best_v3ReasonEval-7B-calibrated-Llama-3.2-1B-InstructMistral_nemo_calibrated_f1enhanced_10shot_oldinstruct_finalv2math-shepherd-mistral-7b-prm-calibrated-Llama-3.1-8B-InstructReasonEval-7B-calibrated-Llama-3.1-8B-InstructQwen2.5-Math-PRM-7B-calibrated-Llama-3.2-1B-InstructMistral_nemo_calibrated_f1enhanced_10shot_oldinstruct_finalqwen2.5-14b-calibrated
