Measured harness ledgerPublic result
Claude Fable 5Explorable space-flight game — Claude Fable 5 Max
Build a responsive browser space-flight game with flight controls, a coherent star-system environment, lighting, assets, and a playable game loop.
Max reasoningHeadline result
- Workflow cost
- $151.55
- Wall-clock
- 1:05:19.2 wall-clock
- Processed tokens
- 92.10M processed
- Record state
- partial_token_timing_and_source_ledger
Public summary
Claude Fable 5 Max partial_token_timing_and_source_ledger ledger: 1:05:19.2 wall-clock, 92.10M processed, and $151.55 API-equivalent estimate, not a subscription invoice.
Run identity and stack
- Result ID: space-flight-game-fable-5-max
- Technical model: claude-fable-5
- Provider: Anthropic Claude Code
- Stack: Anthropic Claude Code
- Stack: Technical model/configuration: claude-fable-5
- Stack: Three.js / Vite game
- Stack: Generated spaceship assets
- Stack: Harness v1 space-flight prompt
Cost basis
- If every cache write used the official one-hour rate instead, the total would be $167.36.
- Primary cost uses 5-minute cache writes.
- All-1-hour cache-write alternative: $167.36.
- No deduplication audit was supplied for this usage receipt; the token and cost figures retain the supplied basis.
Primary artifact integrity
- Kind: interactive-threejs-space-flight-game-entry-point
- Path: artifacts/space-flight-game-fable-5-max/source/src/main.js
- SHA-256: 7fcb9aa3801a71e100a6cc64807308dc23ef4dbf454475c3bebd9c9a7790a94f
Recorded caveats
- Wall-clock is end-to-end latency including idle and tool time, not model-only compute.
- Output tokens include hidden reasoning, code, and tool-call JSON.
- Cache reads are billed at a deep discount, so total processed tokens substantially overstate cost.
- The supplied receipt assumes 5-minute cache writes. It does not separately identify cache-write TTLs.
- No deduplication audit was supplied with this metrics receipt; the token and cost figures retain the supplied basis.
- The build passed, but no local FPS measurement, browser/hardware environment, final capture, or blind evaluation was supplied.
Visible evidence gaps
- browser and hardware environment
- local FPS measurement receipt
- final capture
- blind-evaluation record
Builder test available
This result is part of a Builder test. Open it for the exact prompt and any released projects, RemakeBench Harness workflows and production skills. Public proof and known evidence gaps stay visible here.
- Kimi K3 Launch 002 · v1
