Claude Fable 5.1 in Claude Code 2.1.280
skills off
Zoom, inline player and more ways to view
Wall time 23m 40sCost $6.28 reported
Review
No verdict yet.
The full review
OK. Finished on its own after 23.7 min and produced brown_pelican.svg.
Scores
AI judge
Not scored
No rubric for this prompt
Visitor votes and reviewer grades are not available yet.
- AI judge
- AI judge: the test runner (model claude-opus-5-5), scoring by visual review of the 1600 px render against the pre-registered rubric. No AI-judge score exists for this result: there is no rubric for this prompt, so no score is shown and none is made up.
- Visitors
- Public voting with a blind reveal is coming: you score the output first, then see the other scores.
- Reviewer
- The reviewer's own grade appears here once it is added.
Scores compare results for the same prompt only. AI-judge boards rank runs within one prompt, equal totals sharing a rank; scores from different prompts are never averaged.
Downloads
Files actually published for this run. Original output, site-made previews and post-run conversions are labelled separately.
The exact prompt
agentic harness (save the file; self-render allowed), 561 characters. Every result for this prompt
Generate an SVG of a California brown pelican riding a bicycle. The bicycle must have spokes and a correctly shaped bicycle frame. The pelican must have its characteristic large pouch, and there should be a clear indication of feathers. The pelican must be clearly pedaling the bicycle. The image should show the full breeding plumage of the California brown pelican. Save it as brown_pelican.svg in the current directory. You may render and inspect your own output (headless chromium is available at /usr/bin/chromium) and iterate on your own before finishing.
How it was made
Versions saved during the run
4 intermediate outputs, in the order the model wrote them.
Step by step
The prompt, every agent turn and tool call (tool name and a one-line summary; inputs and outputs are not shown), then the final message. The full transcript is 3.0 MB; it is published as a scrubbed, compressed download rather than shown on the page.
Show all 27 steps
Loading the timeline…
Download the full transcript Claude Code session JSONL, 82 KB, gzip. Scrubbed: anything that looks like a key, token, e-mail address or phone number is replaced, and account details and the test machine's file paths are removed.
Cost breakdown
Headline cost $6.28, as reported by Claude Code (its own bill for the run). The list-price estimate from the token counts is $6.27.
| Tokens by type | Tokens | $ per 1M | Cost |
|---|---|---|---|
| Input (uncached) | 612 | $10.00 | $0.0061 |
| Cache write | 96,423 | $12.50 | $1.2053 |
| Cache read | 980,903 | $0.25 | $0.2452 |
| Output | 96,263 | $50.00 | $4.8132 |
| of which reasoning | not recorded | billed as output |
- Total computed from tokens
- $6.2698
- Total reported by the harness
- $6.2800
- Difference
- computed $6.27 vs reported $6.28 (diff -0.01)
Tokens summed from the Claude Code session JSONL (per message id, incl. subagent files); harness total from /cost.
Prices: Project price table (vendor list prices); cache write 1.25x input, 2026-09-22.
Run notes
- A well-known benchmark author's detailed pelican prompt, verbatim from the image alt text of the post (his own comment in brackets after it left out).
- Mode: interactive TUI in tmux (160x48), inside bubblewrap sandbox
- Isolation: bwrap: private /tmp, HOME and TMPDIR inside run dir, own CLAUDE_CONFIG_DIR/CODEX_HOME copy; the project files, other runs, /tmp/hx_* not mounted




