models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
chinese-alpaca-2-1.3b-rlhf-ggufbloom-560m-RLHF-SD2-prompter-aesthetic-i1-GGUFgpt2-rlhf-implementation-GGUFLlama-3-8b-rlhf-100k-GGUFgpt2-rlhf-anthropic-GGUFbloom-560m-RLHF-SD2-prompter-i1-GGUFkolibri-qwen2.5-7b-060225-rlhf-1-GGUFGPT2-124-poetry-RLHF-GGUFintuitionism-rlhf-GGUFOpenBezoar-HH-RLHF-DPO-GGUFLlama-3.2-1B-Instruct-RLHF-v0.1-GGUFchinese-alpaca-2-7b-rlhf-ggufLinkbricks-Horizon-AI-Korean-llama3.1-sft-rlhf-dpo-8Bbloom-560m-RLHF-SD2-prompter-aesthetic-GGUFMetaAligner-HH-RLHF-7B-i1-GGUFQwen2-7B-Merged-SPPO-Online-RLHF-GGUFbloom-560m-RLHF-SD2-prompter-GGUFDDeduPModelv7-RLHFv2-GGUFMetaAligner-HH-RLHF-1.1B-i1-GGUFdeberta-v3-large-tasksource-rlhf-reward-modelRLHF-VMetaAligner-HH-RLHF-7B-GGUFPythia-2.8B-HH-RLHF-Iterative-SamPO-i1-GGUFRLHF-V-SFTSafe-RLHF-DPO-helpless-mistral-7b-GGUFQwen2-7B-Merged-SPPO-Online-RLHF-i1-GGUFSafe-RLHF-DPO-naive-baseline-llama3-8b-GGUFtulu-v2.5-dpo-13b-hh-rlhf-60kSafe-RLHF-SFT-llama3-3b-GGUFSafe-RLHF-DPO-naive-baseline-llama3-3b-GGUF
