AI Humanizer Benchmark

Leaderboard / Compare / GPTinf vs SuperHumanizer

Head to head · September 2026 cycle

GPTinf vs SuperHumanizer

GPTinf finished 9.8 points ahead of SuperHumanizer in the September 2026 cycle, 78.30 to 68.50, ranking #4 against #8 in a field of 11. GPTinf scored higher on 6 of 7 detectors and 6 of 7 writing categories. SuperHumanizer came out ahead on Meaning preservation (77.3 vs 75.7), Originality.ai (50.8 vs 33.8), News article (85.5 vs 73.2). The bypass-rate gap, 70.7 to 79.9, sits inside both tools' 95% confidence intervals, so treat bypass as a draw.

GPTinfLeads
gptinf.com
78.30/100
Overall score
#4/11
Rank
+9.8
GPTinf leads
Detectors
61
Categories
61
Components
41
SuperHumanizer
superhumanizer.ai
68.50/100
Overall score
#8/11
Rank
Last tested
2026-09-03
Prompts
33
Methodology
v1.0.0

Score components

What the overall score is made of: bypass rate weighs 42%, meaning preservation 32%, readability 16%, consistency across categories 10%. Output-quality penalties come off the total. How scoring works

GPTinfSuperHumanizer
Bypass rateCIs overlapGPTinf by 9.2
79.9
70.7
Meaning preservationSuperHumanizer by 1.6
75.7
77.3
ReadabilityGPTinf by 2.7
76.7
74.0
ConsistencyGPTinf by 9.5
82.2
72.8
Penaltiesfewer is betterGPTinf by 5.0
None
−5.0

CIs overlap marks a gap that sits inside both tools' 95% confidence intervals for bypass rate; it may not survive another cycle.

Detector by detector

Share of each tool's outputs that a detector classified as human-written. GPTinf took 6 of 7.

GPTinfSuperHumanizer
GPTZeroGPTinf by 5.8
69.8
64.0
Winston AIGPTinf by 10.7
63.0
52.3
Originality.aiSuperHumanizer by 17.0
33.8
50.8
ZeroGPTGPTinf by 6.5
85.2
78.7
GrammarlyGPTinf by 10.8
87.6
76.8
QuillBotGPTinf by 10.3
84.9
74.6
CopyleaksGPTinf by 6.2
57.6
51.4

Writing categories

Category score per writing context: bypass credit only for real rewrites, blended with the category's own meaning preservation and readability. Small categories swing; the prompt count is shown on each row.

Academic essayApplication essayBlog postBusiness emailMarketing copyDiscussion boardNews article
GPTinfSuperHumanizerOuter ring = 100
GPTinfSuperHumanizer
Academic essay6 promptsGPTinf by 2.9
85.3
82.4
Application essay6 promptsGPTinf by 15.0
79.5
64.5
Blog post6 promptsGPTinf by 6.9
75.0
68.1
Business email3 promptsGPTinf by 4.7
61.8
57.1
Marketing copy6 promptsGPTinf by 4.1
81.7
77.6
Discussion board3 promptsGPTinf by 1.6
82.9
81.3
News article3 promptsSuperHumanizer by 12.3
73.2
85.5

Where each one wins

Every measure above that one tool won outright, largest margin first within each group.

GPTinf

16 wins
  • Consistency Score82.2 vs 72.8
  • Bypass rate Score79.9 vs 70.7
  • Penalties ScoreNone vs −5.0
  • Readability Score76.7 vs 74.0
  • Grammarly Detector87.6 vs 76.8
  • Winston AI Detector63.0 vs 52.3
  • QuillBot Detector84.9 vs 74.6
  • ZeroGPT Detector85.2 vs 78.7
  • Copyleaks Detector57.6 vs 51.4
  • GPTZero Detector69.8 vs 64.0
  • Application essay Category79.5 vs 64.5
  • Blog post Category75.0 vs 68.1
  • Business email Category61.8 vs 57.1
  • Marketing copy Category81.7 vs 77.6
  • Academic essay Category85.3 vs 82.4
  • Discussion board Category82.9 vs 81.3

SuperHumanizer

3 wins
  • Meaning preservation Score77.3 vs 75.7
  • Originality.ai Detector50.8 vs 33.8
  • News article Category85.5 vs 73.2

Questions people ask

Is GPTinf better than SuperHumanizer?

GPTinf finished 9.8 points ahead of SuperHumanizer in the September 2026 cycle, 78.30 to 68.50, ranking #4 against #8 in a field of 11. GPTinf scored higher on 6 of 7 detectors and 6 of 7 writing categories. Both tools ran the same 33 prompts, scored by the same 7 detectors, under methodology v1.0.0; every input, output, and verdict is published.

Which bypasses AI detectors better, GPTinf or SuperHumanizer?

GPTinf posted the higher bypass rate, 79.9 vs 70.7, and scored higher on 6 of 7 detectors. GPTinf led on GPTZero, Winston AI, ZeroGPT, Grammarly, QuillBot and Copyleaks; SuperHumanizer led on Originality.ai. The overall bypass gap is inside both 95% confidence intervals, so it is not a reliable difference.

Do GPTinf and SuperHumanizer keep the original meaning?

SuperHumanizer preserved meaning better, 77.3 vs 75.7 on our 0–100 similarity scale. On readability, GPTinf rated higher, 76.7 vs 74.0. Penalties this cycle: GPTinf None, SuperHumanizer −5.0.

More matchups

Add a third tool to this comparison·Every matchup

Go deeper