
AI Humanizer Benchmark
Monthly rankings of AI humanizers tested against major detectors with all raw data published for reproducibility.

About AI Humanizer Benchmark
AI Humanizer Benchmark tests major AI humanizers against a panel of popular commercial AI detectors including GPTZero, Originality.ai, and Copyleaks. Each monthly cycle runs identical texts through every tool and measures bypass rate, meaning preservation, readability, and consistency. The site publishes every prompt, raw output, detector verdict, and scoring script on GitHub so anyone can verify the rankings.
Highlights
- Tests 11 AI humanizers against 7 major detectors including GPTZero, Originality.ai, and Copyleaks
- Runs 33 identical texts through each tool monthly and measures bypass rate, meaning preservation, readability, and consistency
- Publishes all prompts, raw outputs, detector verdicts, and scoring scripts on GitHub for reproducibility
- Generated 2,541 detector verdicts in the September 2026 cycle
Who it's for
- Researchers evaluating AI detection evasion techniques
- Content creators comparing AI humanizer tools
- Anyone seeking transparent, reproducible AI tool benchmarks
FAQ
- Which AI humanizers and detectors are tested?
- The benchmark tests 11 AI humanizers against 7 detectors: GPTZero, Originality.ai, Copyleaks, Winston AI, ZeroGPT, QuillBot, and Grammarly.
- How often are the rankings updated?
- Rankings are updated monthly. The most recent cycle tested on September 3, 2026.
- Can I verify the results myself?
- Yes. All prompts, raw outputs, detector verdicts, and the scoring script are published on GitHub so rankings can be reproduced.
- What metrics are measured?
- Each tool is measured on bypass rate, meaning preservation, readability, and consistency across seven writing categories.
At a glance
Sep 2026
Listed since
Category
New tools every Monday
The week's new listings and the most viewed tools, in one email — read the latest. No spam, unsubscribe anytime.