Claude Opus 5.5 in Claude Code 2.1.280
skills off
Zoom, inline player and more ways to view
The model's own page, run in a sandboxed frame only when you ask (a page that needs its own origin opens in a new tab on a separate demo site). Phones get a full-screen player capped at 1.5x pixel density (2x on larger screens) so it stays smooth; the original page runs uncapped. It starts in its auto-tour (T switches to play mode on a keyboard).
Wall time 14m 39sCost $2.46 estimated
Review
No verdict yet.
The full review
OK. Finished on its own after 14.7 min.
Attempt 2 of 2
Every try of this model in this harness on this prompt, oldest first. This run is the cell's result.
- Attempt 1 OK2m 16s$0.442026-09-29
- Attempt 2 OK this run14m 39s$2.462026-09-30
Scores
AI judge
Not scored
No rubric for this prompt
Visitor votes and reviewer grades are not available yet.
- AI judge
- AI judge: the test runner (model claude-opus-5-5), scoring by visual review of the 1600 px render against the pre-registered rubric. No AI-judge score exists for this result: there is no rubric for this prompt, so no score is shown and none is made up.
- Visitors
- Public voting with a blind reveal is coming: you score the output first, then see the other scores.
- Reviewer
- The reviewer's own grade appears here once it is added.
Scores compare results for the same prompt only. AI-judge boards rank runs within one prompt, equal totals sharing a rank; scores from different prompts are never averaged.
Downloads
Files actually published for this run. Original output, site-made previews and post-run conversions are labelled separately.
The exact prompt
as run, 492 characters. Every result for this prompt
Build a complete, polished Battleship game playable in the browser (index.html plus any JS/CSS, no build step). Player vs computer on 10x10 grids: ship placement (drag or click, with rotate), a computer opponent with a sensible hunt/target strategy, hit/miss/sunk feedback with animation and sound, turn indicator, win/lose screen and restart. It must work with mouse and touch and look good on desktop and phone. Write everything into the working directory and finish with a short README.md.
How it was made
Versions saved during the run
2 intermediate outputs, in the order the model wrote them.
Step by step
The prompt, every agent turn and tool call (tool name and a one-line summary; inputs and outputs are not shown), then the final message. The full machine-format transcript is published as a scrubbed download.
Show all 31 steps
Loading the timeline…
Download the full transcript Claude Code session JSONL, 489 KB. Scrubbed: anything that looks like a key, token, e-mail address or phone number is replaced, and account details and the test machine's file paths are removed.
Cost breakdown
Headline cost $2.46, estimated from the token counts at list price.
| Tokens by type | Tokens | $ per 1M | Cost |
|---|---|---|---|
| Input (uncached) | 36 | $4.00 | $0.0001 |
| Cache write | 96,815 | $5.00 | $0.4841 |
| Cache read | 1,629,841 | $0.20 | $0.3260 |
| Output | 82,325 | $20.00 | $1.6465 |
| of which reasoning | 37,686 | billed as output |
- Total computed from tokens
- $2.4567
- Total reported by the harness
- none
Computed from harness-reported token counts x list prices.
Prices: Project price table (vendor list prices), 2026-09-22.
Run notes
- frozen prompt (written), used verbatim
- plain wave-1-equivalent cc_config copy (settings.json + .claude.json; no ECC files)
- Mode: interactive TUI in tmux (160x48), inside bubblewrap sandbox
- Context setting: 1M ([1m] model name)
- Isolation audit: none


