models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
ReSearch-HotpotQA-PRMSearch-o1-HotpotQA-PRMReAct-HotpotQA-PRMReAct-MedQA-PRMmath-shepherd-mistral-7b-prm-calibrated-Llama-3.2-1B-InstructQwen2.5-Math-PRM-7B-calibrated-Llama-3.2-1B-Instructmath-shepherd-mistral-7b-prm-calibrated-Qwen2.5-Math-1.5B-InstructQwen2.5-Math-PRM-7B-calibrated-Qwen2.5-Math-1.5B-InstructPRO-STEP-PRM-8Bprm800k_llama_loraprm800k_qwen_alt_loraSearch-o1-MedQA-PRMprm800k_qwen_loraDirect-HotpotQA-PRMthinktank-prm-qwen2.5-0.5bnpc-fin-prm-7bcode-prm-critic-lora-32kQwen3.5-27B-prm-ep1prm800k-ds_qwen_loraCoT-HotpotQA-PRMDirect-MedQA-PRMprm-qwen3-8b-bf16-6kQwen2.5-Math-PRM-7B-calibrated-Qwen2.5-Math-7B-InstructQwen3.5-27B-SFT-SuccessFiltered-PRMGuidedReSearch-MedQA-PRMRAG-MedQA-PRMcodellama-7b-openapi-completion-ctx-lvl-prmtCoT-MedQA-PRMmath-shepherd-mistral-7b-prm-calibrated-DeepSeek-R1-Distill-Llama-8BQwen2.5-7B-PRM_lora_CL_1
