GPT-6 Astra in OpenCode 1.18.32
skills off
Zoom, inline player and more ways to view
Wall time 7m 08sCost $0.94 reported9.0/10 · Draft score — AI judge
Review
Same house style as Astra in Codex (cream vignette, teal bike, scarf), with a nicely drawn bike and a believable pose. The bird leans stork-like (long neck, black wing feathers) more than pelican.
Verdict by the AI judge (provisional).
The full review
OK. Finished on its own after 7.1 min and produced pelican.svg.
Bugs
- Black-tipped wing and long neck read more stork than pelican
- Minimal scene
Scores
AI judge
9.0/10
Draft score — AI judge
Visitor votes and reviewer grades are not available yet.
| Criterion | Marks |
|---|---|
| pelican | 1.5 |
| bicycle | 2 |
| riding | 2 |
| composition | 1.5 |
| craft | 2 |
| Total (sum of the marks) | 9 / 10 |
- AI judge
- AI judge: the test runner (model claude-opus-5-5), scoring by visual review of the 1600 px render against the pre-registered rubric. A draft score until the reviewer checks it. Rubric: Wave-1 SVG scoring rubric (AI Model Reviewer).
- Visitors
- Public voting with a blind reveal is coming: you score the output first, then see the other scores.
- Reviewer
- The reviewer's own grade appears here once it is added.
Scores compare results for the same prompt only. AI-judge boards rank runs within one prompt, equal totals sharing a rank; scores from different prompts are never averaged.
Downloads
Files actually published for this run. Original output, site-made previews and post-run conversions are labelled separately.
The exact prompt
agentic harness (save the file; self-render allowed), 234 characters. Every result for this prompt
Generate an SVG of a pelican riding a bicycle. Save it as pelican.svg in the current directory. You may render and inspect your own output (headless chromium is available at /usr/bin/chromium) and iterate on your own before finishing.
How it was made
Versions saved during the run
2 intermediate outputs, in the order the model wrote them.
Step by step
The prompt, every agent turn and tool call (tool name and a one-line summary; inputs and outputs are not shown), then the final message. The full machine-format transcript is published as a scrubbed download.
Show all 18 steps
Loading the timeline…
Download the full transcript OpenCode session export, 95 KB. Scrubbed: anything that looks like a key, token, e-mail address or phone number is replaced, and account details and the test machine's file paths are removed.
Cost breakdown
Headline cost $0.94, as reported by OpenCode (its own bill for the run). The list-price estimate from the token counts is $0.89.
| Tokens by type | Tokens | $ per 1M | Cost |
|---|---|---|---|
| Input (uncached) | 20,917 | $10.00 | $0.2092 |
| Cache read | 128,678 | $1.00 | $0.1287 |
| Output | 10,982 | $50.00 | $0.5491 |
| of which reasoning | 3,829 | billed as output |
- Total computed from tokens
- $0.8869
- Total reported by the harness
- $0.9392
- Difference
- computed 0.8869 vs harness 0.9392
Tokens from OpenCode per-message usage; harness total = sum of OpenCode per-message cost. OpenCode reports reasoning separately; both bill at the output rate (output line here = output only, see reasoning).
Prices: Project price table; OpenAI Standard tier, requests whose prompt is >272K tokens bill at 2x input/cache and 1.5x output (applied per request where per-request usage is logged), 2026-09-22.
Run notes
- a well-known benchmark wording "Generate an SVG of a pelican riding a bicycle" (a full stop added before the harness suffix).
- per-request tiering: requests with a prompt above 272K tokens billed at 2x input/cache, 1.5x output
- Mode: interactive TUI in tmux (160x48), inside bubblewrap sandbox
- Isolation: bwrap: private /tmp and /run, private PID ns; only this run dir, harness binaries (ro), opencode config, /usr, /etc bound; HOME/TMPDIR/XDG data private; the project files and other runs not mounted


