datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
best-ai-humanizer-independent-benchmark
AI Humanizer Benchmark: Rankings & Public Audit Record
The best AI humanizers, ranked by a monthly benchmark. This benchmark measures how well each AI humanizer bypasses the major AI detectors (GPTZero, Originality.ai, Copyleaks, Winston AI, and ZeroGPT) while preserving the original meaning and readability. We pay for every tool ourselves and run each one by hand on the most undetectable setting it advertises; there are no affiliate deals and no vendor-supplied numbers. Every… See the full description on the dataset page: https://huggingface.co/datasets/AIHumanizerBenchmarks/best-ai-humanizer-independent-benchmark.ai-humanizer-benchmark
AI Humanizer Benchmark — monthly cycle data
The complete raw data of AI Humanizer Benchmark, a monthly measured benchmark of AI humanizers. Every tool rewrites the same 33 freshly generated texts on its default settings; every output is scored by 7 commercial AI detectors (GPTZero, Originality.ai, Copyleaks, Winston AI, ZeroGPT, QuillBot, Grammarly) plus meaning preservation and readability.
This dataset is the official mirror of the GitHub data repository, published by the AI… See the full description on the dataset page: https://huggingface.co/datasets/ai-humanizer-benchmark/ai-humanizer-benchmark.ai-humanizer
AI Humanizer Dataset (JSONL)
This dataset is designed for fine-tuning instruction-following LLMs
to rewrite AI-generated text into more natural, human-like language.
Structure
train.jsonl – training split
validation.jsonl – validation split
Format
Each line is a JSON object:
{
"prompt": "Rewrite the following text to sound natural, human-like, and conversational...",
"completion": "Humanized output text here",
"attribution": "Original… See the full description on the dataset page: https://huggingface.co/datasets/KNipun/ai-humanizer.ai-humanizer-benchmark-2026
AI Humanizer Benchmark 2026 (Rephrasy)
Every detector result behind the Best AI Humanizer 2026 ranking, one row per (tool, text, detector, run). Nine humanizers, two detectors, four test batches from December 2025 to September 2026.
Nothing here is aggregated into a number you cannot trace. Each row has a source column: a public blog post with screenshots, or a screenshot in the org-card assets folder.
Files
benchmark.csv – 30 rows. Columns: test_id, date, tool… See the full description on the dataset page: https://huggingface.co/datasets/ai-humanizer/ai-humanizer-benchmark-2026.
