September 2026 review · same 50 texts
Grubby vs RewriteAI
In the September 2026 HumanizerEval review, Grubby and RewriteAI rewrote the same 50 English texts. Grubby cleared all 5 detectors on 18 of them; RewriteAI on 13. The difference is mostly ZeroGPT: on 6 samples Grubby passed it where RewriteAI failed, and the reverse never happened. On Winston AI the two were close to even.
Per-sample agreement
| Detector | Passed more often | Margin | Both passed | Neither passed |
|---|---|---|---|---|
| Pangram 4.0 | Grubby | 10 to 6 | 8 | 26 |
| GPTZero 2026-09-13-base | Grubby | 6 to 2 | 41 | 1 |
| Originality.ai turbo | Grubby | 6 to 1 | 43 | 0 |
| ZeroGPT default | Grubby | 6 to 0 | 44 | 0 |
| Winston AI 5.0 | Grubby | 2 to 0 | 48 | 0 |
| All 5 detectors | Grubby | 10 to 5 | 8 | 27 |
Every count is computed from scores.csv at build time, not transcribed. “Neither passed” counts the texts that defeated both tools; on Pangram that was 26 of 50 here.
Quality and output discipline
| Measure | Grubby | RewriteAI |
|---|---|---|
| 72% | 54% | |
| 46% | 40% | |
| 98% | 100% | |
| 94% | 94% | |
| 1.11 | 1.09 | |
| 0 | 2 | |
| 0 | 0 | |
| Academic | API default, passages of up to 280 words (300-word API limit) | |
| Website | grubby.ai | rewriteai.com |
On per-sample quality scores, Grubby scored higher on 23 samples, RewriteAI on 12, and the two tied on 15. RewriteAI’s output was shorter on 29 of 50 samples; mean length ratios were 1.11 and 1.09, or 11% longer than the input and 9% longer than the input respectively.
Where each one outperforms the other
Grubby
- Pangram: passed 18 of 50 against 14, including 10 samples where RewriteAI failed, against 6 the other way.
- GPTZero: passed 47 of 50 against 43, including 6 samples where RewriteAI failed, against 2 the other way.
- Originality.ai: passed 49 of 50 against 44, including 6 samples where RewriteAI failed, against 1 the other way.
- ZeroGPT: passed 50 of 50 against 44, including 6 samples where RewriteAI failed, against 0 the other way.
- Winston AI: passed 50 of 50 against 48, including 2 samples where RewriteAI failed, against 0 the other way.
- Factual consistency: 72% against 54%.
- Naturalness: 46% against 40%.
RewriteAI
- Pangram: passed 6 samples where Grubby failed, even though it was detected more often overall, passing 14 to 18.
- GPTZero: passed 2 samples where Grubby failed, even though it was detected more often overall, passing 43 to 47.
- Originality.ai: passed 1 sample where Grubby failed, even though it was detected more often overall, passing 44 to 49.
- Syntax integrity: 100% against 98%.
- Length: shorter output on 29 of 50 samples, which matters under a word limit.
Grubby vs RewriteAI FAQ
Is Grubby better than RewriteAI?
In the September 2026 HumanizerEval review, Grubby cleared all 5 detectors on 18 of the 50 shared samples and RewriteAI on 13. On Winston AI the two were within 6 points of each other. Quality ran 23 samples to 12, with 15 ties.
Which humanizer is best for Pangram, Grubby or RewriteAI?
Grubby. In September 2026, Grubby passed Pangram on 18 of the 50 shared samples and RewriteAI on 14. On 10 samples Grubby passed where RewriteAI failed; the reverse happened on 6. Both failed on 26.
Which humanizer changes the text less?
Grubby. It scored higher on per-sample quality on 23 of 50 samples, against 12. Grubby’s outputs averaged 1.11× the input length and passed factual consistency on 72%; RewriteAI’s averaged 1.09× and passed on 54%. On stance preservation the split was 94% to 94%. RewriteAI produced the shorter output on 29 of 50 samples.
More comparisons
- StealthGPT vs Grubby
- StealthGPT vs RewriteAI
- WriteHuman vs Grubby
- WriteHuman vs RewriteAI
- Grubby vs Rephrasy
- Grubby vs Undetectable.ai
- Grubby vs HIX Bypass
- Grubby vs Humbot
- Grubby vs BypassGPT
- Grubby vs Ryne
- Rephrasy vs RewriteAI
- Undetectable.ai vs RewriteAI
- RewriteAI vs HIX Bypass
- RewriteAI vs Humbot
- RewriteAI vs BypassGPT
- RewriteAI vs Ryne
- Full September 2026 leaderboard: all 10 tools, 96.8% down to 23.6%