September 2026 review · rank 3 of 10
Grubby AI detector test results
In the September 2026 HumanizerEval review, Grubby ranked 3 of 10 with a mean strict-bypass rate of 85.6%, behind StealthGPT Super at 96.8%. It cleared all 5 detectors on 18 of 50 texts. Review our full results for the best AI humanizer on the leaderboard page.
Grubby results, September 2026
- Website
- grubby.ai
- Rank
- 3 of 10
- 85.6%
- 18 of 50
- 18 of 50 (36%), mean 39.8%, median 10.69%
- 47 of 50 (94%), mean 93.6%, median 100%
- 49 of 50 (98%), mean 98.2%, median 100%
- 50 of 50 (100%), mean 96.0%, median 100%
- 50 of 50 (100%), mean 99.3%, median 99.85%
- 72%
- 46%
- 98%
- 94%
- 1.11
- 0
- 0
- Academic
Recompute any of this from scores.csv. Strict bypass means an overall human verdict on a human score above 50%.
What the vendor claims, and what we measured
- Claim
- “Our bypass ai tool helps you achieve 99%+ human scores on even the most advanced AI detectors, including GPTZero, Originality.ai, ZeroGPT, Turnitin, Winston AI, Content at Scale, and Copyleaks.” (grubby.ai , retrieved October 1, 2026)
- Measured
- Strict bypass in September 2026: Pangram 36%, GPTZero 94%, Originality.ai 98%, ZeroGPT 100%, Winston AI 100%.
- Verdict
- Not verified. In September 2026, strict bypass was GPTZero 94% against the 99% claimed, Originality.ai 98% against the 99% claimed, ZeroGPT 100% against the 99% claimed and Winston AI 100% against the 99% claimed. Turnitin, Content at Scale and Copyleaks are not in this panel.
Review history
| Review | Rank | Mean bypass | Pangram | GPTZero | Originality.ai | ZeroGPT | Winston AI | Naturalness |
|---|---|---|---|---|---|---|---|---|
| September 2026 | 3 of 10 | 85.6% | 36% | 94% | 98% | 100% | 100% | 46% |
Head to head
Because every tool saw the same texts, these pages report per-sample agreement, how often each tool bypassed detection on the same input, rather than a difference of averages.
Grubby vs StealthGPT
Grubby cleared all 5 detectors on 1 sample where StealthGPT did not; the reverse happened on 29. Both cleared 17.
Grubby vs WriteHuman
Grubby cleared all 5 detectors on 2 samples where WriteHuman did not; the reverse happened on 17. Both cleared 15.
Grubby vs Rephrasy
Grubby cleared all 5 detectors on 7 samples where Rephrasy did not; the reverse happened on 9. Both cleared 11.
Grubby vs Undetectable.ai
Grubby cleared all 5 detectors on 9 samples where Undetectable.ai did not; the reverse happened on 6. Both cleared 9.
Grubby vs RewriteAI
Grubby cleared all 5 detectors on 10 samples where RewriteAI did not; the reverse happened on 5. Both cleared 8.
Grubby vs HIX Bypass
Grubby cleared all 5 detectors on 17 samples where HIX Bypass did not; the reverse never happened. Both cleared 1.
Grubby vs Humbot
Grubby cleared all 5 detectors on 18 samples where Humbot did not; the reverse never happened. Both cleared 0.
Grubby vs BypassGPT
Grubby cleared all 5 detectors on 18 samples where BypassGPT did not; the reverse never happened. Both cleared 0.
Grubby vs Ryne
Grubby cleared all 5 detectors on 18 samples where Ryne did not; the reverse never happened. Both cleared 0.
Grubby FAQ
Does Grubby work?
No, it does not reliably bypass Pangram. In the September 2026 HumanizerEval review it passed Pangram on 36%, GPTZero on 94%, Originality.ai on 98%, ZeroGPT on 100%, Winston AI on 100% of texts. Its mean bypass was 85.6%, rank 3 of 10.
Does Grubby change the meaning of the text?
Yes. In September 2026, 72% of its outputs passed the factual-consistency check and 94% passed stance preservation.
Does Grubby make text longer?
Yes. Its outputs averaged 11% longer than the input in September 2026 (ratio 1.11). No sample crossed the 1.4× inflation or 0.6× deflation threshold.