Head-to-head on the Towards AI writing benchmark: same 10 real YouTube scripts, five runs each, scored blind by a three-family judge panel. Claude Fable 5 (max) leads overall, 89.3 to 86.8.
Fable 5 wins this one, and the intervals say it's real: 2320.9 Elo at #1 against GLM-5.3's 2041.5 at #24, and the confidence bands don't come close to touching. Overall it's 89.3 to 86.83, a 2.47-point gap. GLM-5.3 actually takes one metric, hook strength, 89.54 to 89.39, but Fable wins everything else: length adherence is the widest gap at 93.2 against 87.5, YouTube best practices follows at 89.0 to 85.02, and continuity and emotion goes 88.26 to 85.11. The part the averages hide is consistency. Fable's overall spread is 1.55, GLM's is 6.6, so GLM swings from strong scripts to rough ones while Fable barely moves. The records tell the same story, 1185 wins and a single loss against 866 wins and 8 losses. Now the case for GLM: it's open weights, about eleven cents per script against roughly twenty-six for Fable, and it's noticeably faster. Staying within 2.47 points of the board leader while being self-hostable at under half the price is a real result. The inconsistency is what would worry me on anything that ships without an editing pass.
Pick Claude Fable 5 (max effort + 4.8 fallback) if you want the #1 board result with a tight 1.55-point spread, so every script lands close to its 89.3 average at roughly twenty-six cents each.
Pick GLM-5.3 if you want open weights you can self-host at about eleven cents per script and can live with output that swings between excellent and rough.
Blue bars: Claude Fable 5 (max). Orange bars: GLM-5.3. Same 0–100 scale; the bold bar wins that metric.
| Claude Fable 5 (max) | GLM-5.3 | |
|---|---|---|
| Overall / 100 | 89.3 | 86.8 |
| Writing Elo | 2321 | 2042 |
| Run-to-run spread (± overall std) | 1.550 | 6.600 |
| Cost per script (USD) | 0.256 | 0.109 |
| Avg latency (s) | 546.9 | 301.9 |
| Open weights | No | Yes |
Full scorecards: Claude Fable 5 (max) · GLM-5.3. How scoring works: methodology.
← All comparisons