Measured harness ledgerPublic result
Claude Fable 5Campfire under the stars — Claude Fable 5 Max
Build a polished interactive Three.js campfire scene under a starry night with convincing fire, lighting, environmental detail, and camera motion.
Max reasoningHeadline result
- Workflow cost
- $11.17
- Wall-clock
- 12:48.9 wall-clock
- Processed tokens
- 2.95M processed
- Record state
- partial_token_timing_and_source_ledger
Public summary
Claude Fable 5 Max partial_token_timing_and_source_ledger ledger: 12:48.9 wall-clock, 2.95M processed, and $11.17 API-equivalent estimate, not a subscription invoice.
Run identity and stack
- Result ID: campfire-threejs-fable-5-max
- Technical model: claude-fable-5
- Provider: Anthropic Claude Code
- Stack: Anthropic Claude Code
- Stack: Technical model/configuration: claude-fable-5
- Stack: Three.js interactive scene
- Stack: Harness v1 campfire prompt
Cost basis
- The supplied receipt applies standard Anthropic cache-rate multipliers because cache rates were not itemized on the fetched official page.
Primary artifact integrity
- Kind: interactive-threejs-campfire-scene
- Path: artifacts/campfire-threejs-fable-5-max/source/main.js
- SHA-256: 811cbf890d3da4a66042c62d262c523a338e04f95edd2ffc7c43f98469c14d14
Recorded caveats
- Wall-clock is end-to-end latency, including tool execution, idle gaps, and human think time; it is not model-only compute.
- Output tokens include hidden reasoning, code, and tool-call JSON.
- Cache reads are billed at a deep discount, so total processed tokens overstate cost.
- The exact v1 script overcounted because Claude Code repeats a message usage object per content block. This ledger uses the supplied deduplicated unique-message basis instead.
- The raw and deduplicated metrics snapshots were taken minutes apart while the transcript was still growing; the supplied deduplicated basis is the best API-billing estimate.
- The source was archived without a fresh browser run in this environment; no local FPS, browser/hardware environment, final capture, or blind evaluation was supplied.
Visible evidence gaps
- browser and hardware environment
- local FPS measurement receipt
- final capture
- blind-evaluation record
Builder test available
This result is part of a Builder test. Open it for the exact prompt and any released projects, RemakeBench Harness workflows and production skills. Public proof and known evidence gaps stay visible here.
- Kimi K3 Launch 002 · v1
