models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
ArmoRM-Llama3-8B-v0.1llama-3-tulu-2-8b-uf-mean-rmFsfairX-LLaMA3-RM-v0.1Llama-3.1-8B-Instruct-RM-RB2llama-3.1-8b-oracle-rm-hh-rlhf-harmlessnessllama-3.1-8b-oracle-rm-hh-rlhf-helpfulnessLlama-3.1-Tulu-3-8B-RMLlama-3-OffsetBias-RM-8BLLaMA-3-8B-SFR-RM-RLlama-3.1-8B-Energy-ClassifierLlama-3.1-8B-SemiEvol-MMLUINF-ORM-Llama3.1-70BLlama-3.1-8B-Base-RM-RB2llama-3-tulu-2-70b-uf-mean-rmLlama-3.1-70B-Instruct-RM-RB2Llama-3.1-Tulu-3-8B-DPO-RM-RB2google-long-t5-tglobal-base_finetuned_augmented_augmented_llama3.3_70bLlama-3.1-Tulu-3-8B-RL-RM-RB2Llama-3.1-Tulu-3-8B-SFT-RM-RB2Llama-3.1-Tulu-3-70B-SFT-RM-RB2Llama-3.1-8B-Instruct-RobloxGuard-1.0GRM-llama3-8B-distillGRM-llama3-8B-sftregllama-3.2-qlora-safety-classifierLlama-3-8B-ORPOllama3.1-8b-classifier-josh-dsllama-3-8b-Instruct-bnb-4bit-sentiment_100_try9Llama-3.2-1B-imdbLlama-3.2-1B-Instruct_Bradly-Terry-RM_Preference-700kroberta-llama3.1405B-twitter-sentiment
