BenchLeader

LMArena Creative Writing

Text arena rating on creative-writing prompts.

As of 19 Sept 2026, Claude Fable 5 leads LMArena Creative Writing on BenchLeader with 1504, ahead of Claude Opus 4.6 at 1500, across 364 model configurations with a published result.

Published by
LMArena
Category
Human preference
Index weight
Reference only
Models
364
Data as of
19 Sept 2026

CC BY 4.0 — lmarena-ai/leaderboard-dataset on Hugging Face.

What the test looks like

Pairwise votes on stories, poems and other creative prompts.

How it is scored

Bradley-Terry rating from pairwise votes, in Elo-like units, with style control.

What to keep in mind

Taste-driven by design; the style control removes length and formatting effects but not voice preferences.

364 of 364
#
1Claude Fable 5Anthropic150468.32026-09-13
2Claude Opus 4.6highAnthropic150061.32026-09-13
3Gemini 3.8 FlashhighGoogle149664.52026-09-13
4Gemini 3.7 FlashhighGoogle149564.62026-09-13
5Claude Opus 4.7highAnthropic148862.42026-09-13
6Claude Fable 5.1maxAnthropic148669.72026-09-13
7Gemini 3 ProGoogle148361.12026-09-13
8Claude Opus 4.7Anthropic148164.52026-09-13
9Gemini 3.1 ProGoogle148063.92026-09-13
10Claude Opus 4.6Anthropic147963.72026-09-13
11Claude Opus 5highAnthropic147470.22026-09-13
12GPT-5.6 SolxhighOpenAI147368.22026-09-13
13Claude Opus 5maxAnthropic147269.92026-09-13
14Qwen3 8maxAlibaba147266.52026-09-13
15Gemini 3.6 FlashhighGoogle147061.72026-09-13
16Claude Opus 4.5highAnthropic146857.72026-09-13
17Gemini 3.5 FlashmediumGoogle146664.22026-09-13
18Gemini 3.5 FlashhighGoogle146563.62026-09-13
19Claude Opus 4.8highAnthropic146562.22026-09-13
20Muse SparkMeta146465.72026-09-13
21Claude Opus 4.8Anthropic146361.92026-09-13
22GLM-5.3maxZhipu AIopen ↗146265.82026-09-13
23Grok 4.20 Beta1xAI14622026-09-13
24GPT-6 AstramaxOpenAI146171.82026-09-13
25Claude Opus 4.5Anthropic146158.22026-09-13
26Gemini 3 FlashGoogle145957.72026-09-13
27Kimi K3maxMoonshot AIopen ↗145867.22026-09-13
28GPT-5.5 InstantOpenAI145659.22026-09-13
29Claude Sonnet 4.5Anthropic145354.32026-09-13
30Muse Spark 1.3maxMeta145369.32026-09-13
31Muse Spark 1.2xhighMeta145264.12026-09-13
32Grok 4.6highxAI145164.52026-09-13
33GPT-5.5highOpenAI145167.02026-09-13
34Muse Spark 1.1Meta145165.22026-09-13
35GLM-5.2maxZhipu AIopen ↗145163.92026-09-13
36Grok 4.5xAI145060.32026-09-13
37Claude Sonnet 4.6Anthropic145059.02026-09-13
38Claude Sonnet 4.5highAnthropic144958.82026-09-13
39GPT-5.5OpenAI144863.22026-09-13
40Grok 4.20 Multi-AgentxAI144859.22026-09-13
41GLM-5Zhipu AIopen144854.72026-09-13
42Qwen3 5maxAlibaba14472026-09-13
43GLM-5.1Zhipu AIopen ↗144756.92026-09-13
44Grok 4.20thinkingxAI144659.92026-09-13
45Gemini 3 FlashminimalGoogle144654.92026-09-13
46Claude Opus 4.1thinkingAnthropic144557.22026-09-13
47GPT-5.4highOpenAI144459.02026-09-13
48DeepSeek V4 ProDeepSeekopen ↗144452.62026-09-13
49Qwen3 7maxAlibaba144461.02026-09-13
50Gemini 2.5 ProGoogle144354.52026-09-13
51Deepseek v4 ProhighDeepSeek144260.82026-09-13
52Claude Opus 4.1Anthropic144251.92026-09-13
53Qwen3 6maxAlibaba143863.32026-09-13
54Claude Sonnet 5highAnthropic143762.22026-09-13
55GPT-5.4OpenAI143759.22026-09-13
56GPT-4.5OpenAI143650.22026-09-13
57Gemini 3.5 Flash LiteGoogle143655.52026-09-13
58MiMo-V2.5-ProXiaomiopen ↗143559.22026-09-13
59GLM-5.3-FlashZhipu AIopen ↗143563.92026-09-13
60GPT-5.2OpenAI143358.42026-09-13

Cite as: BenchLeader, “LMArena Creative Writing leaderboard”, https://www.benchleader.com/benchmarks/lmarena_text_creative, data as of 19 Sept 2026.

LMArena Creative Writing: questions

What does LMArena Creative Writing measure?
Pairwise votes on stories, poems and other creative prompts. Scores are reported in Elo-style ratings from pairwise comparisons; higher is better.
Which AI model leads LMArena Creative Writing?
Claude Fable 5 leads LMArena Creative Writing with 1504 as of 19 Sept 2026, ahead of Claude Opus 4.6 at 1500.
How many models have LMArena Creative Writing results?
364 model configurations have a LMArena Creative Writing result on BenchLeader, all taken from LMArena.
Who runs LMArena Creative Writing and how often is it updated?
LMArena Creative Writing is published by LMArena. BenchLeader re-reads the published results every morning and records the date each result was published.
Does LMArena Creative Writing count toward the BenchLeader Index?
No. LMArena Creative Writing is shown for reference but left out of the composite index.