datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
LLM-EvaluationHub
LLM-EvaluationHub: Enhanced Dataset for Large Language Model Assessment
This repository, LLM-EvaluationHub, presents an enhanced dataset tailored for the evaluation and assessment of Large Language Models (LLMs). It builds upon the dataset originally provided by SafetyBench (THU-COAI), incorporating significant modifications and additions to address specific research objectives. Below is a summary of the key differences and enhancements:
Key Modifications… See the full description on the dataset page: https://huggingface.co/datasets/strikoder/LLM-EvaluationHub.LLM_EVAL_Datasetllm_evaluationllm-eval-recipe-impact
