datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Salesforce__LLaMA-3-8B-SFR-Iterative-DPO-R-details
Dataset Card for Evaluation run of Salesforce/LLaMA-3-8B-SFR-Iterative-DPO-R
Dataset automatically created during the evaluation run of model Salesforce/LLaMA-3-8B-SFR-Iterative-DPO-R
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/Salesforce__LLaMA-3-8B-SFR-Iterative-DPO-R-details.shared-imagination
Dataset Card for Shared Imagination
This dataset contains the problems used in the paper Shared
Dataset Description
This dataset contains the questions generated for the investigations described in the TMLR paper Shared Imagination: LLMs Hallucinate Alike.
If you want to use this dataset to assess new models, please use the default config (i.e., datasets.load_dataset('Salesforce/shared-imagination')).
This config contains questions for which the four candidate choices… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/shared-imagination.InstruSum
InstruSum
This is the dataset corresponding to our paper "Benchmarking Generation and Evaluation Capabilities of Large Language
Models for Instruction Controllable Summarization".
dataset
The dataset subset contains 100 human-written data examples by us.
Each example contains an article, a summary instruction, a LLM-generated summary, and a hybrid LLM-human summary.
human_eval
This subset contains human evaluation results for the 100 examples in the dataset… See the full description on the dataset page: https://huggingface.co/datasets/Salesforce/InstruSum.
