Recommendation for Creative & fiction

Creative Writing

Our top recommendation for Creative Writing, based on the public evidence we track, is Anthropic: Claude Opus 4.6.[1][2] Use when you need top-tier creative writing judged by human taste, as it scores second-highest among 146 models on LMArena's creative-writing category. Z.ai: GLM 5.2 is the next-ranked alternative. Use as an open-weight option for hard science fiction physics consultation, as one author employed it alongside Gemini, Claude, and OpenAI models to review story physics.

About this recommendation

Updated
Sep 24, 2026
Evidence through
Sep 24, 2026
Sources
2
Revision
v76

Decision audit

Why this result

Inspect the inputs and the computed order behind the recommendation.

Models screened

22

live candidates

Evaluation feeds

5

task-weighted

Winner coverage

100%

intended feed weight

Largest provider share

1 of 2

Anthropic

Provisional source breadth. 2 citation families and 0 practitioner families support the top result; 0 cautionary threads is retained. The largest citation family contributes 50%.

Sources evaluated

The task sets these weights before any model is scored.

winner: Claude Opus 4.6
Evaluation feedWeightWinner resultField measured
LMArena Creative Writing
30%
#222/22
LiveBench Instruction Following
25%
#3822/22
LMArena Text
20%
#222/22
LiveBench Language
15%
#1622/22
OpenRouter usage
10%
83/10022/22

Provider concentration

Each exact model is scored separately; provider identity is not a ranking input.

Anthropic50%
  • Anthropic1 model
  • Z.ai1 model

Decision table

Every published model is shown in computed order. Practitioner sources are distinct community threads, not the citations repeated in the prose below.

RankModelRelative scoreCoveragePractitioner evidenceStrongest measured reason
01Claude Opus 4.6Anthropic
79
100%no linked practitioner threads#2 LMArena Creative Writing · #2 LMArena Text
02GLM 5.2Z.ai
73
100%1 threads · 1 families · 1 cautions#12 LMArena Creative Writing · #24 LMArena Text

Relative score combines normalized benchmark quality and signal coverage; independent practitioner evidence and freshness are bounded tie-breakers. It is an ordering score, not an absolute quality percentage. The writing model receives this order and cannot change it.

  1. Claude Opus 4.6 sits at #2 on LMArena's creative-writing leaderboard, indicating strong human preference for its prose quality and voice in blind evaluations.

    Best when: Use when you need top-tier creative writing judged by human taste, as it scores second-highest among 146 models on LMArena's creative-writing category.

    Tips

    • Use when you need top-tier creative writing judged by human taste, as it scores second-highest among 146 models on LMArena's creative-writing category.
      Source 1
      “Ranks #2 of 146 on LMArena's creative-writing category (Elo 1505), based on blind human preference votes.”
      LMArena creative-writing categoryOpen original ↗
  2. GLM 5.2 is an open-weight model with no direct creative-writing benchmark evidence, though it has been used in multi-model physics debates for hard science fiction.

    Best when: Use as an open-weight option for hard science fiction physics consultation, as one author employed it alongside Gemini, Claude, and OpenAI models to review story physics.

    Tips

    • Use as an open-weight option for hard science fiction physics consultation, as one author employed it alongside Gemini, Claude, and OpenAI models to review story physics.
      Source 2
      “> it is clear that actual intelligence has plateaued significantly N=1, but I disagree strongly. I'm writing a hard-science science fiction story, and the physics of it is at (and frankly, beyond) my skillset. The story's plot has had to change over a dozen times as I realized errors in my application of physics in the story. Throughout, I've been reviewing the physics with LLMs, mainly Gemini 3.1 Pro Preview, but also with Claude and OpenAI. Often I have the LLMs debate each other -- "My frien…”

Frequently asked

What is the top-ranked model for Creative Writing?
Anthropic: Claude Opus 4.6 ranks first in the current evidence-weighted comparison. Use when you need top-tier creative writing judged by human taste, as it scores second-highest among 146 models on LMArena's creative-writing category.[1]
What is an alternative to Anthropic: Claude Opus 4.6?
Z.ai: GLM 5.2 is the next-ranked option. Use as an open-weight option for hard science fiction physics consultation, as one author employed it alongside Gemini, Claude, and OpenAI models to review story physics.[2]

Sources

  1. 1

    “Ranks #2 of 146 on LMArena's creative-writing category (Elo 1505), based on blind human preference votes.”

    LMArena creative-writing category · Benchmark · Sep 13, 2026
  2. 2

    “> it is clear that actual intelligence has plateaued significantly N=1, but I disagree strongly. I'm writing a hard-science science fiction story, and the physics of it is at (and frankly, beyond) my skillset. The story's plot has had to change over a dozen times as I realized errors in my application of physics in the story. Throughout, I've been reviewing the physics with LLMs, mainly Gemini 3.1 Pro Preview, but also with Claude and OpenAI. Often I have the LLMs debate each other -- "My frien…”

    gcanyon · Hacker News · Jun 20, 2026

Rankings synthesized from community evidence and open benchmarks. See methodology. Not driven by vendor marketing.