AI Humanizer Benchmark

Detectors / QuillBot

Best for QuillBot

Humanizers ranked by their measured bypass rate against QuillBot. This cycle it was the 6th-strictest of the 7 detectors in the panel: humanizers averaged 73.6% bypass against it.

A free detector inside the QuillBot writing suite, widely used by students.

Reports AI percentage; inverted and divided by 100.

Last tested
2026-09-03T16:11:14.775Z
Sample size
33
Methodology
v1.0.0
AI humanizer ranking for the September 2026 cycle, ordered by QuillBot bypass.
RankHumanizerPenaltiesLast tested
1SmartHumanizersmarthumanizer.ai94.284.80nonetoday
2UndetectedGPTundetectedgpt.ai92.784.90nonetoday
3WriteHumanwritehuman.ai92.281.60-1.0today
4AI Humanizeaihumanize.io89.376.10nonetoday
5HIX Bypassbypass.hix.ai88.973.80nonetoday
6CleverHumanizercleverhumanizer.ai88.866.80-11.0today
7GPTinfgptinf.com84.978.30nonetoday
8SuperHumanizersuperhumanizer.ai74.668.50-5.0today
9StealthGPTstealthgpt.ai66.071.90-3.0today
10ReHumanizerehumanize.io21.939.00-4.0today
11Undetectable AIundetectable.ai16.638.20-12.0today

11 humanizers · 7 detectors · scores out of 100, higher is better. Click a column to sort; hover a penalty for the breakdown.

Raw test data

Every verdict QuillBot returned in the September 2026 cycle as JSONL: the text scored and the score, one row per test. Paste any row's text back into the detector to check the recorded verdict.

QuillBot
363 rows · .jsonl

How scores are computed

See the full methodology

Overall score formula

42%
Bypass

How often the rewritten text is classified as human by the detector panel.

32%
Meaning

How closely the rewrite preserves the original text's meaning.

16%
Readability

Whether the output is clear, fluent, natural writing.

10%
Consistency

Whether performance holds across all seven writing types or varies widely between them.

Penalties (right) subtract from the composite, and the result never goes below zero.

Possible penalties

CodePenaltyMax
  • Identical to input

    The text came back essentially unchanged; no meaningful rewriting occurred.

    -2.0
    -10.0
  • Refusal

    The tool declined to process the text, typically due to its own content filter.

    -1.0
    -10.0
  • Meaning drift

    The rewrite departed substantially from the original's meaning.

    -1.0
    -10.0
  • Length inflation

    The output ran far longer than the input, diluting the AI signal with added words rather than removing it.

    -1.0
    -10.0
  • Length deflation

    The output ran far shorter than the input; content was dropped rather than rephrased.

    -1.0
    -10.0

Maximum possible total penalty: -50.0 (off 100).