CoolFace
7 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01open-llm-leaderboard /databricks__dolly-v2-7b-detailsgated Dataset Card for Evaluation run of databricks/dolly-v2-7b Dataset automatically created during the evaluation run of model databricks/dolly-v2-7b The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/databricks__dolly-v2-7b-details.tabular10K<n<100K0 likes59 downloads2y agoHugging Face02open-llm-leaderboard /databricks__dolly-v2-12b-detailsgated Dataset Card for Evaluation run of databricks/dolly-v2-12b Dataset automatically created during the evaluation run of model databricks/dolly-v2-12b The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/databricks__dolly-v2-12b-details.tabular10K<n<100K0 likes44 downloads2y agoHugging Face03open-llm-leaderboard /databricks__dolly-v1-6b-detailsgated Dataset Card for Evaluation run of databricks/dolly-v1-6b Dataset automatically created during the evaluation run of model databricks/dolly-v1-6b The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/databricks__dolly-v1-6b-details.tabular10K<n<100K0 likes23 downloads2y agoHugging Face04open-llm-leaderboard /databricks__dolly-v2-3b-detailsgated Dataset Card for Evaluation run of databricks/dolly-v2-3b Dataset automatically created during the evaluation run of model databricks/dolly-v2-3b The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/databricks__dolly-v2-3b-details.tabular10K<n<100K0 likes22 downloads2y agoHugging Face05Inversta /rationale-databricks-dolly-cqa Dataset Overview Filtered and annotated version of the closed-question answering part (~1.5k datapoints) of the Databricks Dolly Dataset intended for the task of rationale extraction. Citation @article{pirenne2024exploration, title={Exploration of Closed-Domain Question Answering Explainability Methods With a Sentence-Level Rationale Dataset}, author={Pirenne, Lize and Mokeddem, Samy and Ernst, Damien and Louppe, Gilles}, year={2024} }… See the full description on the dataset page: https://huggingface.co/datasets/Inversta/rationale-databricks-dolly-cqa.tabularquestion-answering1K<n<10K1 likes22 downloads2y agoHugging Face06dataformer /dolly-llama-qa Dataset Card for dolly-llama-qa This dataset has been created with dataformer. Dataset Details Dataset Description The dolly-llama-qa dataset is a synthetic QA pair dataset created using the context from databricks-dolly-15k. We used Meta-Llama-3-8B-Instruct and Meta-Llama-3.1-8B-Instruct models for the generation and evolution part. Openai's gpt-4o was used for evaluating the refined questions and refined answers. Dataset Columns context:… See the full description on the dataset page: https://huggingface.co/datasets/dataformer/dolly-llama-qa.tabulartext-generation1K<n<10K1 likes11 downloads2y agoHugging Face07hongli-zhan /SPRI-SFT-dollygated Dataset Card for SPRI-SFT-dolly (ICML 2025) Paper: SPRI: Aligning Large Language Models with Context-Situated Principles (Published in ICML 2025) Authors: Hongli Zhan, Muneeza Azmat, Raya Horesh, Junyi Jessy Li, Mikhail Yurochkin Shared by: Hongli Zhan Arxiv Link: arxiv.org/abs/2502.03397 Citation If you used our dataset, please cite our paper: @inproceedings{zhan2025spri, title = {SPRI: Aligning Large Language Models with Context-Situated Principles}, author… See the full description on the dataset page: https://huggingface.co/datasets/hongli-zhan/SPRI-SFT-dolly.tabulartext-generation10K<n<100K1 likes5 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.