datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ai-generated-text-classification
Dataset Card for "ai-generated-text-classification"
More Information needed
persian-ai-generated-text
📝 Persian AI-Generated Text Dataset
A large-scale collection of 10,546 AI-generated Persian (Farsi) texts produced by 70 different large language models across diverse topics and writing styles. This dataset is designed to support research in AI-generated text detection for the Persian language.
Dataset Summary
Attribute
Value
Language
Persian (Farsi)
Total Samples
10,546
Unique Models
70
API Providers
9 (OpenRouter, NVIDIA, Free Endpoint, HF… See the full description on the dataset page: https://huggingface.co/datasets/ehsantorabi/persian-ai-generated-text.AI_Human_generated_movie_reviews
Dataset Card for Dataset Name
This dataset card aims to be a base template for new datasets. It has been generated using this raw template.
Dataset Details
Dataset Description
The "AI_Human_generated_movie_reviews" dataset consists of 5.23k AI-generated movie reviews alongside 5.23k human-written reviews from the Stanford IMDB dataset. The AI reviews were created using several models, including Gemini 1.5 Pro, GPT-3.5-Turbo, and GPT-4.0-Turbo-Preview, via the… See the full description on the dataset page: https://huggingface.co/datasets/Lyra-stellAI/AI_Human_generated_movie_reviews.cmv_ai_generated_pairs_2024_2025
