datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
human-vs-Ai-generated-datasetai-vs-human-generated-datasetpython-docstring-human-gpt-generated-mixhuman-ai-generated-text
Dataset Card for human-ai-generated-text
This dataset has been created with distilabel.
Dataset Summary
This dataset contains a pipeline.yaml which can be used to reproduce the pipeline that generated it in distilabel using the distilabel CLI:
distilabel pipeline run --config "https://huggingface.co/datasets/ardavey/human-ai-generated-text/raw/main/pipeline.yaml"
or explore the configuration:
distilabel pipeline info --config… See the full description on the dataset page: https://huggingface.co/datasets/ardavey/human-ai-generated-text.flickr30k-pt-br-5k-human-generatedai-vs-human-generated-dataset-sampleflickr30k-pt-br-human-generatedgenerated_chat_0_4M_sharegpt_system_human_gptAI_Human_generated_movie_reviews
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
The "AI_Human_generated_movie_reviews" dataset consists of 5.23k AI-generated movie reviews alongside 5.23k human-written reviews from the Stanford IMDB dataset. The AI reviews were created using several models, including Gemini 1.5 Pro, GPT-3.5-Turbo, and GPT-4.0-Turbo-Preview, via the… See the full description on the dataset page: https://huggingface.co/datasets/Lyra-stellAI/AI_Human_generated_movie_reviews.juraj-juraj-python-docstring-human-gpt-generated-mix
