datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
AI-Generated-vs-Real-Images-Datasets
Dataset Card for "AI-Generated-vs-Real-Images-Datasets"
More Information needed
human-vs-Ai-generated-datasetAI-generated-inpaintings-dataset
Dataset Card for "AI-generated-inpaintings-dataset"
More Information needed
generated-dataset-for-VLMai-vs-human-generated-datasetAI-Generated-vs-Real-Images-Datasets
Dataset Card for "AI-Generated-vs-Real-Images-Datasets"
More Information needed
stocks_demo_react_agent_generated_train_datasetHarmAug_generated_dataset
HarmAug: Effective Data Augmentation for Knowledge Distillation of Safety Guard Models
This dataset contains generated prompts and responses using HarmAug: Effective Data Augmentation for Knowledge Distillation of Safety Guard Models.This dataset is also used for training our HarmAug Guard Model.The unsafe-score is measured by Llama-Guard-3.For rows without responses, the unsafe-score indicates the unsafeness of the prompt.For rows with responses, the unsafe-score indicates the… See the full description on the dataset page: https://huggingface.co/datasets/hbseong/HarmAug_generated_dataset.entity-attribute-sft-dataset-GPT-4.0-generated-v1
Entity Attribute Dataset 50k (GPT-4.0 Generated)
Dataset Summary
The Entity Attribute SFT Dataset (GPT-4.0 Generated) is a machine-generated dataset designed for instruction fine-tuning. It includes detailed product information generated based on the title of each product, aiming to create a structured catalog in JSON format. The dataset encompasses a variety of product categories such as food, home and kitchen, clothing, handicrafts, tools, automotive equipment… See the full description on the dataset page: https://huggingface.co/datasets/fibonacciai/entity-attribute-sft-dataset-GPT-4.0-generated-v1.generated_datasetAI-Generated-vs-Real-Images-Datasets
Dataset Card for "AI-Generated-vs-Real-Images-Datasets"
More Information needed
RARE_output_and_generated_datasetsentity-attribute-sft-dataset-GPT-4.0-generated-v1
Entity Attribute Dataset 50k (GPT-4.0 Generated)
Dataset Summary
The Entity Attribute SFT Dataset (GPT-4.0 Generated) is a machine-generated dataset designed for instruction fine-tuning. It includes detailed product information generated based on the title of each product, aiming to create a structured catalog in JSON format. The dataset encompasses a variety of product categories such as food, home and kitchen, clothing, handicrafts, tools, automotive equipment, and… See the full description on the dataset page: https://huggingface.co/datasets/BaSalam/entity-attribute-sft-dataset-GPT-4.0-generated-v1.hallucinated_answer_generated_dataset_cleanedai-vs-human-generated-dataset-sampleentity-attribute-dataset-GPT-3.5-generated-v1
Entity Attribute Dataset 306k (GPT-3.5 generated)
Dataset Summary
The Entity Attribute Dataset 306k (GPT-3.5 generated) is designed for instruction fine-tuning, specifically for the task of generating structured catalogs in JSON format based on product titles. The dataset includes a diverse range of products from various categories such as food, home and kitchen, clothing, handicrafts, tools, automotive equipment, and more.
Usage
This dataset is intended for… See the full description on the dataset page: https://huggingface.co/datasets/BaSalam/entity-attribute-dataset-GPT-3.5-generated-v1.tokenized_generated_ar_en_th_datasets
Dataset Card for "tokenized_generated_ar_en_th_datasets"
More Information needed
GeneratedDatasetNEWgenerated_datasetfutoshiki_generated_dataset_5x5_6-8GPT_Generated_Dataset_V1gpt2_generated_datasetgenerated_ar_en_th_datasets
Dataset Card for "generated_ar_en_th_datasets"
More Information needed
generated-qa-dataset-3
Dataset Card for "generated-qa-dataset-3"
More Information needed
dataset_generated_by_teacher_meta_llama_Llama_2_13b_hf_00-19-15-04-08-25KB-sLLM-QA-Dataset-GeneratedHarmAug_generated_dataset
HarmAug: Effective Data Augmentation for Knowledge Distillation of Safety Guard Models
This dataset contains generated prompts and responses using HarmAug: Effective Data Augmentation for Knowledge Distillation of Safety Guard Models.This dataset is also used for training our HarmAug Guard Model.The unsafe-score is measured by Llama-Guard-3.For rows without responses, the unsafe-score indicates the unsafeness of the prompt.For rows with responses, the unsafe-score indicates the… See the full description on the dataset page: https://huggingface.co/datasets/AnonHB/HarmAug_generated_dataset.synthetic-error-generated-spelling-correction-dataset-100kqa-dataset-generated-21020
Dataset Card for "qa-dataset-generated-21020"
More Information needed
your-first-generated-dataset
Dataset Card for "your-first-generated-dataset"
More Information needed
