55
runs
214
screens
8
agents
410×
cost spread
Fable 5, med (Claude Code), paper
9 Jun 2026 · 37% квоты pro·5h · 23 min 19 s
Fancy

I thought it would burn more than it did. Overall little distinguishable from Opus, in places even simpler.

Opus 4.7 (Cursor), paper
16 Apr 2026 · $12.30 · 15M · 35 min
Fancy

Twice as expensive as Opus 4.7 at Kilo. Checks itself a lot. The result is slightly better

Opus 4.8, xhigh (Claude Code), paper
28 May 2026 · 53% квоты pro·5h (удвоенной)
Fancy

Burns more; not necessarily better than 4.7. Three screens used half of the already-doubled 5h pro quota.

Opus 5, high (Claude Code), paper
24 Jul 2026 · 27% квоты pro·5h · 20 min 21 s
Fancy

Typical Opus, very similar to previous generations. But the color is unusual — where’s the beige, Opus? :–)

Auto (Cursor), paper
$4.33 · 13.7M · 25 min
Mid

A surprisingly good result for Auto mode. The style choice is like from Opus

DeepSeek 4.1 Flash (Cursor), paper
10 Sept 2026 · $0.12 · 14.1M · 30 min 27 sec
Mid

Slow, partly because the model did not read the images itself and instead reached for the available zai MCP server to analyse them. That would make sense for regular DeepSeek, but this one has vision. Lots of re-checking, as if I had typed /goal make the most fantastic design ever.

GLM 5.3 Max (z.ai coding plan) (Cursor), paper
14 Aug 2026 · 9.7M · 34 min
Mid

Clearly much better than GLM 5.2, which couldn't even use tools properly for me (will rerun it). Spent the first 10 minutes studying images via zai MCP; the model isn't multimodal.

GPT 5.4, xhigh (Cursor), paper
5 Mar 2026 · $2.20 · 4.5M · 11 min
Mid

The agent is different, the model is recognizable. Slightly better than in Codex

Grok 4.5 (Cursor), paper
8 Jul 2026 · $1.23 · 4.2M · 3:30
Mid

Very fast, cheap (50% discount for now), overall fine for the money. Promo is primitive

Grok 4.6, xhigh (Cursor), paper
12 Aug 2026 · $1.48 · 4.9M · 9 min
Mid

Fast, overall decent, very distinctive. Pros: didn't turn the desktop UI into a promo site. Cons: still not quite great — could have put images somewhere

Sonnet 4.5, xhigh (Claude Code), paper
29 Sept 2025 · 25% квоты pro·5h · 8 min
Mid

Substantially closer to the Chinese models and the simpler ones. Neat, but completely neutral, completely simplified

Composer 2 (Cursor), paper
19 Mar 2026 · $0.33 · 1.3M · 5 min
Weak

Very primitive, but also very cheap. Doesn't match the claimed level

GPT 5.5, xhigh (Codex), figma
24 Apr 2026 · 16% квоты plus·5h · 20 min
Weak

In Figma slightly better than in Paper. Lays things out monstrously simply

GPT 5.5, xhigh (Codex), paper
24 Apr 2026 · 26% квоты plus·5h · 15 min
Weak

Very cheap compared to opus, the result is accordingly. The worst result among SOTA models

Grok 4.3 (Cursor), paper
30 Apr 2026 · $1.65 · 3M · 18 min
Weak

Mixed impressions. Mobile is generally fine. Promo is worse

MiMo V2.5 Pro (Cursor), paper
22 Apr 2026 · 4.8M из 4.1B включённых · 4.8M
Weak

Included Cursor model; very cheap in subscription terms, weak on design

Haiku 4.5 (Claude Code), paper
15 Oct 2025 · 6% квоты pro·5h · 3 min
Failed tools

Didn't even cope with the tools. There's not a single reason to use it