AI Humanizer Benchmark

Leaderboard / Compare

Compare AI humanizers

Pick two to four tools and read them side by side, or jump to any pair from the matchup grid. Every number comes from the same September 2026 benchmark run: the same prompts, the same detectors, the same scoring, all of it published.

Pick 2 to 4 tools 3 of 4 selected

Any matchup cell or pair below jumps straight to that comparison.

UndetectedGPT#1 · 84.90SmartHumanizer#2 · 84.80WriteHuman#3 · 81.60
Overall
Leaderboard rank#1#2#3
Overall score84.9084.8081.60
Score componentsBypass 42%, meaning 32%, readability 16%, consistency 10%; penalties come off the total.
Bypass rate95% CI shown under each value86.379.093.585.379.491.290.483.697.1
Meaning preservation84.887.676.2
Readability77.375.772.4
Consistency across categories91.588.386.4
PenaltiesNoneNone1.0
Detector bypass ratesShare of runs each detector classified as human-written.
GPTZero80.071.056.5
Winston AI54.451.082.1
Originality.ai79.278.259.2
ZeroGPT85.182.692.0
Grammarly89.386.994.2
QuillBot92.794.292.2
Copyleaks21.215.263.6
Writing categoriesCategory score: this tool's composite-equivalent within that category alone.
Academic essay6 prompts89.289.886.2
Application essay6 prompts86.283.384.9
Blog post6 prompts77.682.480.6
Business email3 prompts79.775.669.5
Marketing copy6 prompts85.889.286.2
Discussion board3 prompts90.388.383.1
News article3 prompts85.979.379.6
AccessAs tested: every tool runs on its default settings and base plan.
Free tierYesYesYes

Every matchup

Each cell is the row tool's overall score minus the column tool's, September 2026 cycle. Green means the row leads, red the column. Click a cell to load that pair above; the last column is each tool's win–loss record against the rest of the field.

55 pairs · 11 tools

Rankings are recomputed every cycle from published inputs, outputs and detector verdicts, last run September 3, 2026. How scoring works