GLM-5.3 in Claude Code 2.1.280
skills off
This is the unfinished output at the moment the run ended (Harness + model mismatch). See the review below.
Wall time 12m 34sCost $0.31 estimated
Review
No verdict yet.
The full review
Harness + model mismatch. Reclassified 2026-09-24 (published as ok): the run ended on 'API Error: 400 This model does not support image inputs' when GLM-5.3 read its own render (…/bmw_m4_render.png) with Claude Code's Read tool; the artifact was already written but the self-review was cut short (Claude Code x text-only model). Re-run on the text-only relay: bmw_m4_side__claude-code__glm-5.3__attempt2.
Attempt 1 of 2
Every try of this model in this harness on this prompt, oldest first. A later attempt is the cell's result: attempt 2.
- Attempt 1 Harness + model mismatch this run12m 34s$0.312026-09-24
- Attempt 2 OK the cell's result21m 01s$1.022026-09-25
Scores
Not scored: only runs that finished on their own are scored.
Downloads
Files actually published for this run. Original output, site-made previews and post-run conversions are labelled separately.
The exact prompt
agentic harness (save the file; self-render allowed), 441 characters. Every result for this prompt
Create a clean SVG illustration of a BMW M4 Competition in a side view. The image must follow a 4:3 aspect ratio. Use the original factory color of the BMW M4 Competition. Focus on coupe proportions and sporty stance. Use vector shapes without gradients. Save it as bmw_m4.svg in the current directory. You may render and inspect your own output (headless chromium is available at /usr/bin/chromium) and iterate on your own before finishing.
How it was made
Step by step
The prompt, every agent turn and tool call (tool name and a one-line summary; inputs and outputs are not shown), then the final message. The full machine-format transcript is published as a scrubbed download.
Show all 7 steps
Loading the timeline…
Download the full transcript Claude Code session JSONL, 284 KB. Scrubbed: anything that looks like a key, token, e-mail address or phone number is replaced, and account details and the test machine's file paths are removed.
Cost breakdown
Headline cost $0.31, estimated from the token counts at list price.
| Tokens by type | Tokens | $ per 1M | Cost |
|---|---|---|---|
| Input (uncached) | 62,468 | $1.40 | $0.0875 |
| Cache write | 0 | $0.00 | $0 |
| Cache read | 79,802 | $0.26 | $0.0207 |
| Output | 44,780 | $4.40 | $0.1970 |
| of which reasoning | not recorded | billed as output |
- Total computed from tokens
- $0.3052
- Total reported by the harness
- none
Tokens summed from the Claude Code session JSONL (per message id, incl. subagent files); if Claude Code logged zero usage for this unrecognised model, the /cost screen's per-model usage is used instead. Claude Code tags glm-5p3 as unrecognized_model and its own $ figure has costBasis unknown, so it is NOT used: cost = tokens x Fireworks GLM-5.3 price.
Prices: Fireworks GLM-5.3 serverless price (STATE.md 22:42Z; same as the OpenCode/models.dev catalog entry fireworks-ai/glm-5p3), 2026-09-23.
Run notes
- Verbatim; the post's five lines are joined with spaces so the prompt pastes into every TUI as one message.
- Mode: interactive TUI in tmux (160x48), inside bubblewrap sandbox
- Isolation: bwrap: private /tmp, HOME and TMPDIR inside run dir, own CLAUDE_CONFIG_DIR copy; the project files, other runs, /tmp/hx_* not mounted

