CoolFace
Datasetpublic

OALL/details_RLHFlow__LLaMA3-iterative-DPO-final

Dataset Card for Evaluation run of RLHFlow/LLaMA3-iterative-DPO-final Dataset automatically created during the evaluation run of model RLHFlow/LLaMA3-iterative-DPO-final. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_RLHFlow__LLaMA3-iterative-DPO-final.

sourceHugging Faceupdated 2y agoView on Hugging Face
0likes10downloads
5 commits on main
7e4c7632y ago

Upload README.md with huggingface_hub

Hamza-Alobeidli
24e5bcc2y ago

Upload folder using huggingface_hub

Hamza-Alobeidli
66e27712y ago

Upload results_2024-06-05T19-32-16.175656.parquet with huggingface_hub

Hamza-Alobeidli
8164e052y ago

Upload results_2024-06-05T19-32-16.175656.json with huggingface_hub

Hamza-Alobeidli
76eaa272y ago

initial commit

Hamza-Alobeidli