How it works

One prompt. Two models. You judge the pixels, not the price tag.

Every page in the test is generated cold from a single prompt - one take, n=1, no edits, no cherry-picking. GLM 5.2 (open, on Together AI) and Claude Opus 4.8 (closed) each get the exact same brief, and the page you see is the raw output. The only thing hidden during the test is the bill.

$0.66GLM 5.2 · all 11
$3.01Opus 4.8 · all 11
4.6×cheaper, in total
  1. 1
    Same brief, both models. A short, identical prompt goes to GLM 5.2 and Opus 4.8 - no design kit, no tools, no planning step.
  2. 2
    One raw take each. Whatever the model returns is the page. No retries, no human edits, costed at real API token rates.
  3. 3
    You guess blind. Two unlabelled pages, side by side. Pick the one that cost more - then we show the receipts.

The exact prompt

Identical for both models. Just a system line and a one-paragraph brief - the rest is the model.

System
You are a senior product designer and front-end
engineer. Build a complete, polished landing page
for the brief below as a single, self-contained
HTML file: all CSS and any JS inline, no
frameworks, no build step, no placeholder text.
Make deliberate choices about layout, type,
colour, spacing and motion. Output exactly one
final ```html block.
Brief · per page
Brief: Nimbus - an edge compute platform
that runs your code in 300+ cities with zero
config and instant rollbacks. Audience: backend
and platform engineers. Register: dark mode,
technical, confident, fast.

Build it now. Output one final ```html block.
One of 11 briefs. The category, audience and register change; the shape stays the same.

What each page actually cost

Hover a row to see both pages. Prices are the real API cost of that one generation.

PageGLM 5.2Opus 4.8CheaperRelative cost
NimbusEdge compute platform$0.096$0.303.2×open GLM ↗open Opus ↗
FathomSleep & meditation app$0.060$0.315.1×open GLM ↗open Opus ↗
ForgeStreetwear drop$0.054$0.264.7×open GLM ↗open Opus ↗
Olive & AshFarm-to-table restaurant$0.064$0.284.4×open GLM ↗open Opus ↗
QuantaAI research lab$0.037$0.225.8×open GLM ↗open Opus ↗
VersoDesign magazine$0.048$0.204.1×open GLM ↗open Opus ↗
HalcyonPremium headphones$0.054$0.234.2×open GLM ↗open Opus ↗
MeridianLogistics SaaS$0.068$0.253.7×open GLM ↗open Opus ↗
YonderGuided hiking trips$0.070$0.395.6×open GLM ↗open Opus ↗
SiftEmail client$0.043$0.245.6×open GLM ↗open Opus ↗
BloomPlant subscription$0.062$0.335.4×open GLM ↗open Opus ↗

Methodology

One take per page, n=1, no cherry-picking; both models built from the same prompt and the page you see is the raw output. GLM 5.2 is zai-org/GLM-5.2 on Together AI; Opus is claude-opus-4.8 on the Anthropic API. Prices are the real cost of each generation at list token rates. The per-brief gap runs from 3.2× (Nimbus, narrowest) to 5.8× (Quanta, widest), about 4.6× overall.

Token + time detail
NimbusGLM 22.0k tok · 97sOpus 12.4k tok · 119.11s
FathomGLM 13.8k tok · 57sOpus 12.6k tok · 120.15s
ForgeGLM 12.5k tok · 58sOpus 10.7k tok · 100.4s
Olive & AshGLM 14.7k tok · 53sOpus 11.6k tok · 119.37s
QuantaGLM 8.7k tok · 31sOpus 9.1k tok · 82.64s
VersoGLM 11.1k tok · 38sOpus 8.2k tok · 77.37s
HalcyonGLM 12.6k tok · 40sOpus 9.5k tok · 88.61s
MeridianGLM 15.6k tok · 72sOpus 10.4k tok · 98.31s
YonderGLM 16.0k tok · 56sOpus 15.9k tok · 156.54s
SiftGLM 10.0k tok · 51sOpus 10.0k tok · 97.34s
BloomGLM 14.3k tok · 58sOpus 13.6k tok · 132.24s