models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
ArmoRM-Llama3-8B-v0.1llama-3-tulu-2-8b-uf-mean-rmGRM-Llama3.2-3B-rewardmodel-ftFsfairX-LLaMA3-RM-v0.1Llama-3.1-8B-Instruct-RM-RB2llama-3.1-8b-oracle-rm-hh-rlhf-harmlessnessllama-3.1-8b-oracle-rm-hh-rlhf-helpfulnessllama-3.1-nemoguard-8b-content-safetyllama-3.1-nemoguard-8b-topic-controlLlama-3.1-Tulu-3-8B-RMLlama-3-OffsetBias-RM-8BGRM_Llama3.1_8B_rewardmodel-ftLLaMA-3-8B-SFR-RM-RLlama-3.1-8B-Energy-ClassifierLlama-3.1-8B-SemiEvol-MMLUINF-ORM-Llama3.1-70BLlama-3.1-8B-Base-RM-RB2llama-3-tulu-2-70b-uf-mean-rmLlama-3.1-70B-Instruct-RM-RB2Llama-3.1-Tulu-3-8B-DPO-RM-RB2google-long-t5-tglobal-base_finetuned_augmented_augmented_llama3.3_70bLlama-3.1-Tulu-3-8B-RL-RM-RB2Llama-3.1-Tulu-3-8B-SFT-RM-RB2Llama-3.1-Tulu-3-70B-SFT-RM-RB2Llama-3.1-8B-Instruct-RobloxGuard-1.0GRM-llama3-8B-distillGRM-llama3-8B-sftregllama-3.2-qlora-safety-classifierLlama-3-8B-ORPOAgentDoG-Llama3.1-8B
