datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
panda-bench
PandaBench
PandaBench is a comprehensive benchmark for evaluating Large Language Model (LLM) safety, focusing on jailbreak attacks, defense mechanisms, and evaluation methodologies.
The PandaGuard framework architecture illustrating the end-to-end pipeline for LLM safety evaluation. The system connects three key components: Attackers, Defenders, and Judges.
Dataset Description
This repository contains the benchmark results from extensive evaluations of various… See the full description on the dataset page: https://huggingface.co/datasets/Beijing-AISI/panda-bench.panda-70m
Panda 70M dataset by Snap Inc
70M video-caption pairs
Code for downloading: https://github.com/snap-research/Panda-70M/dataset_dataloading
Panda-70Marm-asmpandas_table_qa_ft_v3arm-asm-xsmallsmart-home-datasetpandas_data_analysis_questionsPandaomicspandas_table_qa_ft_v1text-to-pandaspanda2m_splitPandas-Query-GenerationOCI_pandas_DATASETmanual_datasetV2OCI_pandas_generic_DATASETOCI_pandas_generic_evol-It_DATASETOCI_pandas_generic_self-It_DATASETtext-to-pandas-queries-llmnatural_language_pandas_basketballRAG12000-LLaMA3.1-8B-gguf_AR-RAG_v2ukiyoe-sourceragas_evaluationV1Dataset-responsesv2with-evaluationmerged_response_cprRAG12000-LLaMA3.1-8B-ggufPanda-2M1213
