Measured harness ledgerPublic result
Claude Fable 5.1

The Last of Us Main Menu — Claude Fable 5.1 Max

Recreate the pinned The Last of Us-style main menu reference as an interactive browser experience.

Max reasoningHeadline result
Workflow cost
$31.60
Wall-clock
81m (rounded provider display) wall-clock
Processed tokens
Not recorded
Record state
partial_artifact_static_serve_and_rounded_usage_ledger
Public summary

Claude Fable 5.1 Max partial_artifact_static_serve_and_rounded_usage_ledger ledger: 81m (rounded provider display) wall-clock, Not recorded, and $31.60 provider-reported session cost; not independently recomputable from rounded display values.

Run identity and stack
  • Result ID: last-of-us-main-menu-browser-fable-5.1-max
  • Technical model: claude-fable-5.1
  • Provider: Anthropic Claude Code
  • Stack: Anthropic Claude Code
  • Stack: Technical model/configuration: claude-fable-5.1
  • Stack: Browser recreation
  • Stack: Pinned visual reference
  • Stack: Harness v1 menu prompt
Cost basis
  • No raw session export, exact token counts, cache TTL, or per-line price receipt was supplied.
Primary artifact integrity
  • Kind: interactive-browser-title-menu
  • Path: artifacts/last-of-us-main-menu-browser-fable-5.1-max/source/index.html
  • SHA-256: 71bce8cd53bd08c2c8b274abb0ac943094c20a2064d76234ac1dca2dbdc17444
Recorded caveats
  • Wall-clock and API duration are whole-minute values displayed by the provider, not raw timestamps.
  • The displayed input, cache, and output values are abbreviated; the recorded 79,669,000 processed-token total is an approximation from those displayed components.
  • The $31.60 total is provider-reported and is not a verified API-equivalent cost calculation.
  • The supplied project did not include the exact prompt submission payload, so byte-for-byte candidate prompt-delivery fidelity is unverified.
  • The supplied report does not establish one-shot protocol compliance or session-specific assistant/model-call counts.
  • Rolling local-activity requests are excluded because they are not session-attributed benchmark metrics.
  • Included PNGs are model-supplied project captures, not independent RemakeBench browser replay captures.
  • Install/syntax/static-serve success does not establish runtime interaction, WebGL visual fidelity, audio, local FPS, RTX capture, or blind-vote outcomes.
  • The supplied usage report does not provide user-turn, assistant-turn, or raw transcript evidence.
Visible evidence gaps
  • raw Claude Code session export or exact machine-readable token receipt
  • per-line pricing and cache-TTL receipt sufficient to recompute cost
  • exact candidate prompt submission payload
  • recorded one-shot protocol evidence
  • browser interaction and keyboard-navigation replay receipt
  • local FPS with browser, hardware, and viewport receipt
  • RTX final capture
  • blind-evaluation record
Public result only

This result keeps its public summary and evidence, but it does not currently have a matching Builder test with prompts, projects, Harness workflows or skills.

RemakeBenchResearch console