September 2026 review · same 50 texts
Grubby vs Ryne
In the September 2026 HumanizerEval review, Grubby and Ryne rewrote the same 50 English texts. Grubby cleared all 5 detectors on 18 of them; Ryne on 0. The difference is mostly Originality.ai: on 49 samples Grubby passed it where Ryne failed, and the reverse never happened.
Per-sample agreement
| Detector | Passed more often | Margin | Both passed | Neither passed |
|---|---|---|---|---|
| Pangram 4.0 | Grubby | 11 to 4 | 7 | 28 |
| GPTZero 2026-09-13-base | Grubby | 43 to 0 | 4 | 3 |
| Originality.ai turbo | Grubby | 49 to 0 | 0 | 1 |
| ZeroGPT default | Grubby | 19 to 0 | 31 | 0 |
| Winston AI 5.0 | Grubby | 37 to 0 | 13 | 0 |
| All 5 detectors | Grubby | 18 to 0 | 0 | 32 |
Every count is computed from scores.csv at build time, not transcribed. “Neither passed” counts the texts that defeated both tools; on Pangram that was 28 of 50 here.
Quality and output discipline
| Measure | Grubby | Ryne |
|---|---|---|
| 72% | 44% | |
| 46% | 80% | |
| 98% | 98% | |
| 94% | 94% | |
| 1.11 | 1.07 | |
| 0 | 0 | |
| 0 | 0 | |
| Academic | Aggressive, beast mode | |
| Website | grubby.ai | ryne.ai |
On per-sample quality scores, Grubby scored higher on 14 samples, Ryne on 18, and the two tied on 18. Ryne’s output was shorter on 33 of 50 samples; mean length ratios were 1.11 and 1.07, or 11% longer than the input and 7% longer than the input respectively.
Where each one outperforms the other
Grubby
- Pangram: passed 18 of 50 against 11, including 11 samples where Ryne failed, against 4 the other way.
- GPTZero: passed 47 of 50 against 4, including 43 samples where Ryne failed, against 0 the other way.
- Originality.ai: passed 49 of 50 against 0, including 49 samples where Ryne failed, against 0 the other way.
- ZeroGPT: passed 50 of 50 against 31, including 19 samples where Ryne failed, against 0 the other way.
- Winston AI: passed 50 of 50 against 13, including 37 samples where Ryne failed, against 0 the other way.
- Factual consistency: 72% against 44%.
Ryne
- Pangram: passed 4 samples where Grubby failed, even though it was detected more often overall, passing 11 to 18.
- Naturalness: 80% against 46%.
- Length: shorter output on 33 of 50 samples, which matters under a word limit.
Grubby vs Ryne FAQ
Is Grubby better than Ryne?
In the September 2026 HumanizerEval review, Grubby cleared all 5 detectors on 18 of the 50 shared samples and Ryne on 0. Quality ran 14 samples to 18, with 18 ties.
Which humanizer is best for Pangram, Grubby or Ryne?
Grubby. In September 2026, Grubby passed Pangram on 18 of the 50 shared samples and Ryne on 11. On 11 samples Grubby passed where Ryne failed; the reverse happened on 4. Both failed on 28.
Which humanizer changes the text less?
Ryne. It scored higher on per-sample quality on 18 of 50 samples, against 14. Grubby’s outputs averaged 1.11× the input length and passed factual consistency on 72%; Ryne’s averaged 1.07× and passed on 44%. On stance preservation the split was 94% to 94%. Ryne produced the shorter output on 33 of 50 samples.
More comparisons
- StealthGPT vs Grubby
- StealthGPT vs Ryne
- WriteHuman vs Grubby
- WriteHuman vs Ryne
- Grubby vs Rephrasy
- Grubby vs Undetectable.ai
- Grubby vs RewriteAI
- Grubby vs HIX Bypass
- Grubby vs Humbot
- Grubby vs BypassGPT
- Rephrasy vs Ryne
- Undetectable.ai vs Ryne
- RewriteAI vs Ryne
- HIX Bypass vs Ryne
- Humbot vs Ryne
- BypassGPT vs Ryne
- Full September 2026 leaderboard: all 10 tools, 96.8% down to 23.6%