AI Model Reviewer

Claude Opus 5.5 in Claude Code 2.1.280

skills off

← Open in the gallerySame prompt, all models

Zoom, inline player and more ways to view
Poster frame of Claude Opus 5.5's page
Open the original page

The model's own page, run in a sandboxed frame only when you ask (a page that needs its own origin opens in a new tab on a separate demo site). Phones get a full-screen player capped at 1.5x pixel density (2x on larger screens) so it stays smooth; the original page runs uncapped. It starts in its auto-tour (T switches to play mode on a keyboard).

OKProcedural armor suitlow effortNeutral harness

Wall time 4m 03sCost $0.65 estimated

Review

No verdict yet.

The full review

OK. Finished on its own after 4.0 min.

Attempt 2 of 2

Every try of this model in this harness on this prompt, oldest first. This run is the cell's result.

  1. Attempt 1 OK1h 07m$5.382026-09-30
  2. Attempt 2 OK this run4m 03s$0.652026-09-30

Scores

AI judge

Not scored

No rubric for this prompt

Visitor votes and reviewer grades are not available yet.

AI judge
AI judge: the test runner (model claude-opus-5-5), scoring by visual review of the 1600 px render against the pre-registered rubric. No AI-judge score exists for this result: there is no rubric for this prompt, so no score is shown and none is made up.
Visitors
Public voting with a blind reveal is coming: you score the output first, then see the other scores.
Reviewer
The reviewer's own grade appears here once it is added.

Scores compare results for the same prompt only. AI-judge boards rank runs within one prompt, equal totals sharing a rank; scores from different prompts are never averaged.

Downloads

Files actually published for this run. Original output, site-made previews and post-run conversions are labelled separately.

The exact prompt

as run, 927 characters. Every result for this prompt

Build a single self-contained index.html using Three.js (imported from a pinned CDN; no external models, textures or other assets) that shows a full humanoid powered-armour suit built only from code: primitives, extruded shapes, lathes and generated geometry. Aim for 500 or more distinct mesh parts (plates, joints, pistons, vents, bolts, cables, lights) with a coherent, detailed design and good materials and lighting. Include an assemble / explode animation (parts fly in and lock into place, and can be exploded back out), a sci-fi HUD overlay (targeting reticle, part labels that track parts on screen, a system status panel), orbit controls, and buttons to trigger assemble and explode. It must run smoothly and work with touch on mobile. When the suit is built, log the total number of distinct mesh parts to the console as: PART_COUNT <n>. Write everything into the working directory and finish with a short README.md.

How it was made

Wall time 4m 03sFirst artifact at 2m 27s2 saved versions

Versions saved during the run

2 intermediate outputs, in the order the model wrote them.

Version 2 of 2
  1. Version 1 at 2m 30s thumbnailv1 2:30
  2. Version 2 at 2m 45s thumbnailv2 2:45

Step by step

The prompt, every agent turn and tool call (tool name and a one-line summary; inputs and outputs are not shown), then the final message. The full machine-format transcript is published as a scrubbed download.

Show all 16 steps

Loading the timeline…

Download the full transcript Claude Code session JSONL, 269 KB. Scrubbed: anything that looks like a key, token, e-mail address or phone number is replaced, and account details and the test machine's file paths are removed.

Cost breakdown

Headline cost $0.65, estimated from the token counts at list price.

Tokens by typeTokens$ per 1MCost
Input (uncached)26$4.00$0.0001
Cache write29,434$5.00$0.1472
Cache read616,349$0.20$0.1233
Output18,742$20.00$0.3748
of which reasoning1,144billed as output
Total computed from tokens
$0.6454
Total reported by the harness
none

Computed from harness-reported token counts x list prices.

Prices: Project price table (vendor list prices), 2026-09-22.

Run notes

one prompt, agentic harness (self-review allowed)labelled re-run (of threejs_procedural_armor_suit__claude-code__opus-5-5-low)

  • frozen prompt (written): prompts/threejs_procedural_armor_suit.md verbatim (written 2026-09-30 from a public X post 920-part code-built armour suit UI)
  • plain wave-1-equivalent cc_config copy (settings.json + .claude.json; no ECC files)
  • Mode: interactive TUI in tmux (160x48), inside bubblewrap sandbox
  • Context setting: 1M ([1m] model name)
  • Isolation audit: outside paths seen