CoolFace
Agents
Live
Leaderboard
Models
Community
Search
Create
Alerts
Menu
9 results
ModelBest
ModelBest
Search
in
all
models
datasets
apps
agents
people
projects
Models
All models matching “ModelBest”
ZhangYuchi /
modelbest-Qwen-2.5-Omni-7B-SFT-only
any-to-any
transformers
en
2 likes
9 downloads
3mo ago
Hugging Face
ZhangYuchi /
modelbest-Qwen-2.5-Omni-7B-SFT-with-DPO
any-to-any
transformers
en
1 likes
5 downloads
2mo ago
Hugging Face
iulusoy /
model-best
0 likes
4y ago
Hugging Face
MINATOs /
model-bestv2
0 likes
2y ago
Hugging Face
Akki3110 /
model-best
gated
0 likes
2y ago
Hugging Face
witchking999 /
Model_Best
0 likes
10mo ago
Hugging Face
znSkuby /
modelbesterai
0 likes
7mo ago
Hugging Face
renshengjihe /
model-best
0 likes
4mo ago
Hugging Face
Datasets
All datasets matching “ModelBest”
ZhangYuchi /
modelbest-Qwen-2.5-Omni-7B-SFT-with-DPO-dataset_for_DPO
OminiGAIA-DPO-data This dataset contains the final DPO training pairs used to train ZhangYuchi/modelbest-Qwen-2.5-Omni-7B-SFT-with-DPO. The pairs are derived from 7B-native rollouts on OmniGAIA train questions: Roll out the SFT model on answer-hidden train inputs. Audit each rollout with Gemini using the private reference answer and annotated solution. Locate the first erroneous assistant sub-step. Convert the corrected prefix (tau_win) and the original erroneous prefix… See the full description on the dataset page: https://huggingface.co/datasets/ZhangYuchi/modelbest-Qwen-2.5-Omni-7B-SFT-with-DPO-dataset_for_DPO.
reinforcement-learning
1 likes
46 downloads
2mo ago
Hugging Face