Measured harness ledgerPublic result
Claude Fable 5.1The Last of Us Main Menu — Claude Fable 5.1 Max
Recreate the pinned The Last of Us-style main menu reference as an interactive browser experience.
Max reasoningHeadline result
- Workflow cost
- $31.60
- Wall-clock
- 81m (rounded provider display) wall-clock
- Processed tokens
- Not recorded
- Record state
- partial_artifact_static_serve_and_rounded_usage_ledger
Public summary
Claude Fable 5.1 Max partial_artifact_static_serve_and_rounded_usage_ledger ledger: 81m (rounded provider display) wall-clock, Not recorded, and $31.60 provider-reported session cost; not independently recomputable from rounded display values.
Run identity and stack
- Result ID: last-of-us-main-menu-browser-fable-5.1-max
- Technical model: claude-fable-5.1
- Provider: Anthropic Claude Code
- Stack: Anthropic Claude Code
- Stack: Technical model/configuration: claude-fable-5.1
- Stack: Browser recreation
- Stack: Pinned visual reference
- Stack: Harness v1 menu prompt
Cost basis
- No raw session export, exact token counts, cache TTL, or per-line price receipt was supplied.
Primary artifact integrity
- Kind: interactive-browser-title-menu
- Path: artifacts/last-of-us-main-menu-browser-fable-5.1-max/source/index.html
- SHA-256: 71bce8cd53bd08c2c8b274abb0ac943094c20a2064d76234ac1dca2dbdc17444
Recorded caveats
- Wall-clock and API duration are whole-minute values displayed by the provider, not raw timestamps.
- The displayed input, cache, and output values are abbreviated; the recorded 79,669,000 processed-token total is an approximation from those displayed components.
- The $31.60 total is provider-reported and is not a verified API-equivalent cost calculation.
- The supplied project did not include the exact prompt submission payload, so byte-for-byte candidate prompt-delivery fidelity is unverified.
- The supplied report does not establish one-shot protocol compliance or session-specific assistant/model-call counts.
- Rolling local-activity requests are excluded because they are not session-attributed benchmark metrics.
- Included PNGs are model-supplied project captures, not independent RemakeBench browser replay captures.
- Install/syntax/static-serve success does not establish runtime interaction, WebGL visual fidelity, audio, local FPS, RTX capture, or blind-vote outcomes.
- The supplied usage report does not provide user-turn, assistant-turn, or raw transcript evidence.
Visible evidence gaps
- raw Claude Code session export or exact machine-readable token receipt
- per-line pricing and cache-TTL receipt sufficient to recompute cost
- exact candidate prompt submission payload
- recorded one-shot protocol evidence
- browser interaction and keyboard-navigation replay receipt
- local FPS with browser, hardware, and viewport receipt
- RTX final capture
- blind-evaluation record
Public result only
This result keeps its public summary and evidence, but it does not currently have a matching Builder test with prompts, projects, Harness workflows or skills.
