AI Humanizer Benchmark

Detectors / Winston AI

Best for Winston AI

Humanizers ranked by their measured bypass rate against Winston AI. This cycle it was the 3rd-strictest of the 7 detectors in the panel: humanizers averaged 53.2% bypass against it.

A commercial detector scoring documents for AI involvement, plagiarism, and readability.

Reports a 0-100 human score; divided by 100.

Last tested
2026-09-03T16:11:14.775Z
Sample size
33
Methodology
v1.0.0
AI humanizer ranking for the September 2026 cycle, ordered by Winston AI bypass.
RankHumanizerPenaltiesLast tested
1WriteHumanwritehuman.ai82.181.60-1.0today
2AI Humanizeaihumanize.io81.076.10nonetoday
3CleverHumanizercleverhumanizer.ai70.566.80-11.0today
4HIX Bypassbypass.hix.ai69.173.80nonetoday
5GPTinfgptinf.com63.078.30nonetoday
6StealthGPTstealthgpt.ai62.171.90-3.0today
7UndetectedGPTundetectedgpt.ai54.484.90nonetoday
8SuperHumanizersuperhumanizer.ai52.368.50-5.0today
9SmartHumanizersmarthumanizer.ai51.084.80nonetoday
10Undetectable AIundetectable.ai0.138.20-12.0today
11ReHumanizerehumanize.io0.039.00-4.0today

11 humanizers · 7 detectors · scores out of 100, higher is better. Click a column to sort; hover a penalty for the breakdown.

Raw test data

Every verdict Winston AI returned in the September 2026 cycle as JSONL: the text scored and the score, one row per test. Paste any row's text back into the detector to check the recorded verdict.

Winston AI
363 rows · .jsonl

How scores are computed

See the full methodology

Overall score formula

42%
Bypass

How often the rewritten text is classified as human by the detector panel.

32%
Meaning

How closely the rewrite preserves the original text's meaning.

16%
Readability

Whether the output is clear, fluent, natural writing.

10%
Consistency

Whether performance holds across all seven writing types or varies widely between them.

Penalties (right) subtract from the composite, and the result never goes below zero.

Possible penalties

CodePenaltyMax
  • Identical to input

    The text came back essentially unchanged; no meaningful rewriting occurred.

    -2.0
    -10.0
  • Refusal

    The tool declined to process the text, typically due to its own content filter.

    -1.0
    -10.0
  • Meaning drift

    The rewrite departed substantially from the original's meaning.

    -1.0
    -10.0
  • Length inflation

    The output ran far longer than the input, diluting the AI signal with added words rather than removing it.

    -1.0
    -10.0
  • Length deflation

    The output ran far shorter than the input; content was dropped rather than rephrased.

    -1.0
    -10.0

Maximum possible total penalty: -50.0 (off 100).