CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nvidia /HelpSteer2 HelpSteer2: Open-source dataset for training top-performing reward models HelpSteer2 is an open-source Helpfulness Dataset (CC-BY-4.0) that supports aligning models to become more helpful, factually correct and coherent, while being adjustable in terms of the complexity and verbosity of its responses. This dataset has been created in partnership with Scale AI. When used to tune a Llama 3.1 70B Instruct Model, we achieve 94.1% on RewardBench, which makes it the best Reward Model as… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/HelpSteer2.tabular10K<n<100K456 likes152k downloads2y agoHugging Face02nvidia /HelpSteer HelpSteer: Helpfulness SteerLM Dataset HelpSteer is an open-source Helpfulness Dataset (CC-BY-4.0) that supports aligning models to become more helpful, factually correct and coherent, while being adjustable in terms of the complexity and verbosity of its responses. Leveraging this dataset and SteerLM, we train a Llama 2 70B to reach 7.54 on MT Bench, the highest among models trained on open-source datasets based on MT Bench Leaderboard as of 15 Nov 2023. This model is available on… See the full description on the dataset page: https://huggingface.co/datasets/nvidia/HelpSteer.tabular10K<n<100K252 likes2.5k downloads2y agoHugging Face03nlp-with-deeplearning /Ko.HelpSteer원본 데이터셋: nvidia/HelpSteer tabular10K<n<100K1 likes34 downloads3y agoHugging Face04kunishou /HelpSteer-35k-jaNVIDIA が公開している SteerLM 向けのトライアルデータセット HelpSteerを日本語に自動翻訳したデータセットになります。SteerLM でのアライメントをお試ししたい際にご活用下さい。 SteerLM での LLM トレーニング方法については以下の URL を参考にして下さい。 Announcing NVIDIA SteerLM : https://developer.nvidia.com/blog/announcing-steerlm-a-simple-and-practical-technique-to-customize-llms-during-inference NeMo Aligner : https://github.com/NVIDIA/NeMo-Aligner SteerLM training user guide : https://docs.nvidia.com/nemo-framework/user-guide/latest/modelalignment/steerlm.html [参考] SteerLM :… See the full description on the dataset page: https://huggingface.co/datasets/kunishou/HelpSteer-35k-ja.tabular10K<n<100K3 likes30 downloads3y agoHugging Face05H-D-T /HelpSteer2 HelpSteer2: Open-source dataset for training top-performing reward models HelpSteer2 is an open-source Helpfulness Dataset (CC-BY-4.0) that supports aligning models to become more helpful, factually correct and coherent, while being adjustable in terms of the complexity and verbosity of its responses. This dataset has been created in partnership with Scale AI. When used with a Llama 3 70B Base Model, we achieve 88.8% on RewardBench, which makes it the 4th best Reward Model as of… See the full description on the dataset page: https://huggingface.co/datasets/H-D-T/HelpSteer2.tabular10K<n<100K0 likes16 downloads2y agoHugging Face06lovethayo /HelpSteer HelpSteer: Helpfulness SteerLM Dataset HelpSteer is an open-source Helpfulness Dataset (CC-BY-4.0) that supports aligning models to become more helpful, factually correct and coherent, while being adjustable in terms of the complexity and verbosity of its responses. Leveraging this dataset and SteerLM, we train a Llama 2 70B to reach 7.54 on MT Bench, the highest among models trained on open-source datasets based on MT Bench Leaderboard as of 15 Nov 2023. This model is available on… See the full description on the dataset page: https://huggingface.co/datasets/lovethayo/HelpSteer.tabular10K<n<100K0 likes5 downloads5mo agoHugging Face07cheryyunl /helpsteertabular100K<n<1M0 likes3 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.