September 2026 review · same 49 texts

WriteHuman vs Rephrasy

In the September 2026 HumanizerEval review, WriteHuman and Rephrasy rewrote the same 49 English texts. WriteHuman cleared all 5 detectors on 32 of them; Rephrasy on 19. The difference is mostly Pangram: on 12 samples WriteHuman passed it where Rephrasy failed, and the reverse happened on 2. On ZeroGPT and Winston AI the two were close to even.

Per-sample agreement

Per-sample agreement, September 2026. Counts are out of the 49 shared samples, computed at the 50% human-score threshold. Margin counts the samples only one tool passed, higher count first.
DetectorPassed more oftenMarginBoth passedNeither passed
Pangram 4.0WriteHuman12 to 22015
GPTZero 2026-09-13-baseWriteHuman6 to 0430
Originality.ai turboWriteHuman7 to 1410
ZeroGPT defaultTied0 to 0490
Winston AI 5.0Tied0 to 0490
All 5 detectorsWriteHuman14 to 11816

Every count is computed from scores.csv at build time, not transcribed. “Neither passed” counts the texts that defeated both tools; on Pangram that was 15 of 49 here.

Quality and output discipline

Quality and output discipline, September 2026.
MeasureWriteHumanRephrasy
22%36%
49%30%
98%92%
65%60%
1.021.13
03
00
API defaultv4
Websitewritehuman.ai rephrasy.ai

On per-sample quality scores, WriteHuman scored higher on 18 samples, Rephrasy on 15, and the two tied on 16. WriteHuman’s output was shorter on 32 of 49 samples; mean length ratios were 1.02 and 1.13, or 2% longer than the input and 13% longer than the input respectively.

Where each one outperforms the other

WriteHuman

  • Pangram: passed 32 of 49 against 22, including 12 samples where Rephrasy failed, against 2 the other way.
  • GPTZero: passed 49 of 49 against 43, including 6 samples where Rephrasy failed, against 0 the other way.
  • Originality.ai: passed 48 of 49 against 42, including 7 samples where Rephrasy failed, against 1 the other way.
  • ZeroGPT: tied at 49 of 49 but on a higher mean human score, 94.8% to 92.7%. It passed 0 samples where Rephrasy failed.
  • Winston AI: tied at 49 of 49 but on a higher mean human score, 99.8% to 98.8%. It passed 0 samples where Rephrasy failed.
  • Naturalness: 49% against 30%.
  • Syntax integrity: 98% against 92%.
  • Stance preservation: 65% against 60%.
  • Length: shorter output on 32 of 49 samples, which matters under a word limit.

Rephrasy

  • Pangram: passed 2 samples where WriteHuman failed, even though it was detected more often overall, passing 22 to 32.
  • Originality.ai: passed 1 sample where WriteHuman failed, even though it was detected more often overall, passing 42 to 48.
  • Factual consistency: 36% against 22%.

WriteHuman vs Rephrasy FAQ

Is WriteHuman better than Rephrasy?

In the September 2026 HumanizerEval review, WriteHuman cleared all 5 detectors on 32 of the 49 shared samples and Rephrasy on 19. On ZeroGPT and Winston AI the two were within 6 points of each other. Quality ran 18 samples to 15, with 16 ties.

Which humanizer is best for Pangram, WriteHuman or Rephrasy?

WriteHuman. In September 2026, WriteHuman passed Pangram on 32 of the 49 shared samples and Rephrasy on 22. On 12 samples WriteHuman passed where Rephrasy failed; the reverse happened on 2. Both failed on 15.

Which humanizer changes the text less?

WriteHuman. It scored higher on per-sample quality on 18 of 49 samples, against 15. WriteHuman’s outputs averaged 1.02× the input length and passed factual consistency on 22%; Rephrasy’s averaged 1.13× and passed on 36%. On stance preservation the split was 65% to 60%. WriteHuman produced the shorter output on 32 of 49 samples.

More comparisons