Recommendation for Marketing & copy
Marketing Copy
Our top recommendation for Marketing Copy, based on the public evidence we track, is Anthropic: Claude Fable 5. Anthropic: Claude Opus 4.6 is the next-ranked alternative.
About this recommendation
- Updated
- Sep 4, 2026
- Evidence through
- Sep 4, 2026
- Sources
- 2
- Revision
- v57
Decision audit
Why this result
Inspect the inputs and the computed order behind the recommendation.
Models screened
22
live candidates
Evaluation feeds
5
task-weighted
Winner coverage
100%
intended feed weight
Largest provider share
2 of 4
Anthropic
Sources evaluated
The task sets these weights before any model is scored.
| Evaluation feed | Weight | Winner result | Field measured |
|---|---|---|---|
| LiveBench Instruction Following | 30% | #5 | 21/22 |
| LMArena Instruction Following | 25% | #3 | 21/22 |
| LMArena Text | 20% | #1 | 21/22 |
| LMArena Creative Writing | 15% | #2 | 21/22 |
| OpenRouter usage | 10% | 87/100 | 22/22 |
Provider concentration
Each exact model is scored separately; provider identity is not a ranking input.
- Anthropic2 models
- Google1 model
- Qwen1 model
Decision table
Every published model is shown in computed order. Practitioner sources are distinct community threads, not the citations repeated in the prose below.
| Rank | Model | Relative score | Coverage | Practitioner evidence | Strongest measured reason |
|---|---|---|---|---|---|
| 01 | Claude Fable 5Anthropic | 85 | 100% | no linked practitioner threads | #1 LMArena Text · #2 LMArena Creative Writing |
| 02 | Claude Opus 4.6Anthropic | 81 | 100% | no linked practitioner threads | #1 LMArena Creative Writing · #1 LMArena Instruction Following |
| 03 | Gemini 3.6 FlashGoogle | 80 | 100% | no linked practitioner threads | #7 LiveBench Instruction Following · #10 LMArena Creative Writing |
| 04 | Qwen3.7 MaxQwen | 76 | 100% | no linked practitioner threads | #10 LiveBench Instruction Following · #16 LMArena Text |
Relative score combines normalized benchmark quality and signal coverage; independent practitioner evidence and freshness are bounded tie-breakers. It is an ordering score, not an absolute quality percentage. The writing model receives this order and cannot change it.
Anthropic: Claude Fable 5 ranks #1 of 144 on LMArena's overall text arena (Elo 1507), based on blind human preference votes.
Best when: Consider only after reviewing the cited caution.
Anthropic: Claude Opus 4.6 ranks #2 of 144 on LMArena's overall text arena (Elo 1505), based on blind human preference votes.
Best when: Consider only after reviewing the cited caution.
Achieves #7 on LiveBench instruction following at 75.37% despite a mid-tier LMArena ranking, suggesting particular strength in structured generation tasks over open-ended preference.
Best when: Consider only after reviewing the cited caution.
Watch out for
- Expect lower blind human preference scores compared to top-ranked models, ranking #14 on LMArena's overall text arena.
Ranks #10 on LiveBench instruction following at 74.04% with mid-tier LMArena placement, showing solid task execution without top-tier preference appeal.
Best when: Consider only after reviewing the cited caution.
Watch out for
- Anticipate less compelling open-ended generation compared to leaders, as its #20 LMArena ranking trails top models by 30+ Elo points.
Sources
- 1
“Ranks #14 of 144 on LMArena's overall text arena (Elo 1480), based on blind human preference votes.”
LMArena text arena · Benchmark · Sep 1, 2026 - 2
“Ranks #20 of 144 on LMArena's overall text arena (Elo 1474), based on blind human preference votes.”
LMArena text arena · Benchmark · Sep 1, 2026
Rankings synthesized from community evidence and open benchmarks. See methodology. Not driven by vendor marketing.