models
Open weights, fine-tunes and adapters. Every listing here comes live from the Hugging Face Hub, attributed to it, and links back to the source.
SearchR1-nq_hotpotqa_train-qwen2.5-7b-em-grpo-v0.3SearchR1-nq_hotpotqa_train-qwen2.5-7b-em-ppo-v0.3SearchR1-nq_hotpotqa_train-qwen2.5-7b-it-em-ppohotpotqa_abstractiveSearchR1-nq_hotpotqa_train-qwen2.5-7b-em-ppoSearchR1-nq_hotpotqa_train-qwen2.5-7b-em-ppo-i1-GGUFReSearch-HotpotQA-PRMSearchR1-nq_hotpotqa_train-qwen2.5-3b-it-em-ppoSearch-o1-HotpotQA-PRMSearchR1-nq_hotpotqa_train-qwen2.5-3b-it-em-grpo-v0.3ReAct-HotpotQA-PRMSearchR1-nq_hotpotqa_train-llama3.2-3b-it-em-ppoSearchR1-nq_hotpotqa_train-qwen2.5-7b-em-ppo-GGUFHyLaR-HotpotQA-Qwen2.5-0.5B-Instruct-GGUFSearchR1-nq_hotpotqa_train-qwen2.5-3b-em-ppoHotpotQA-Reader-CoT-Llama-3-8B-InstructHotpotQA-Paragraph-Retriever-Llama-3-8B-InstructHotpotQA-OneStep-Retriever-Llama-3-8B-InstructHotpotQA-Reader-Llama-3-8B-InstructHotpotQA-Sentence-Retriever-Llama-3-8B-Instructhotpotqa_extractive_compressortriviaqa_hotpotqa_train-search-r1-ppo-qwen2.5-7b-em-iter1SearchR1-nq_hotpotqa_train-qwen2.5-3b-em-ppo-v0.3Qwen2.5-3B-UFO-hotpotqa-GGUFQwen2.5-3B-UFO-hotpotqa-1turn-GGUFHotpotQA-Reader-Llama-3-70B-Instructflan_t5_large-kilt_tasks_hotpotqa_final_examSearchR1-nq_hotpotqa_train-qwen2.5-7b-em-grpo-v0.2SearchR1-nq_hotpotqa_train-qwen2.5-3b-it-em-grpocontriever-gpl-hotpotqa
