Measured harness ledgerPublic result
Claude Fable 5.1Space Flight Stage 2 — Landfall — Claude Fable 5.1 Max
Extend the Stage 1 space-flight project with a seamless orbit-to-surface round trip, explorable terrain, takeoff, landing, and a fixed character asset using one frozen follow-on request.
Max reasoningHeadline result
- Workflow cost
- $186.71
- Wall-clock
- 2h 07m 34.4s wall-clock
- Processed tokens
- 124.11M processed
- Record state
- partial_token_timing_source_build_and_two_cycle_browser_replay_ledger
Public summary
Claude Fable 5.1 Max partial_token_timing_source_build_and_two_cycle_browser_replay_ledger ledger: 2h 07m 34.4s wall-clock, 124.11M processed, and $186.71 Official Anthropic API-list-price equivalent, not an itemized Claude Code subscription charge..
Run identity and stack
- Result ID: space-flight-game-threejs-stage-2-landfall-fable-5.1-max
- Technical model: claude-fable-5-1
- Provider: Anthropic Claude Code
- Stack: Anthropic Claude Code
- Stack: Technical model/configuration: claude-fable-5-1
- Stack: Three.js Stage 2 workflow
- Stack: Harness v1 Landfall prompt
Cost basis
- All-5-minute cache-write alternative: $165.64.
Primary artifact integrity
- Kind: manifest-verified-artifact
- Path: artifacts/space-flight-game-threejs-stage-2-landfall-fable-5.1-max/source/src/main.js
- SHA-256: 3a4ea3ad6a0f34ba8ed9db4a4f1e664f27f23c949979c7e6532ebb211ce3913a
Validation evidence
- Result: PASS_WITH_LIMITATIONS
- Path: artifacts/space-flight-game-threejs-stage-2-landfall-fable-5.1-max/validation/build-and-landfall-replay.public.json
- SHA-256: 03af474a04eb8ff6e2fc8ecc38bb6a7c28b0fafa66b9c9f5066b30c798e3c747
Recorded caveats
- Wall-clock is end-to-end workflow latency, including tool execution, browser automation, builds, and user gaps; it is not model-only compute time.
- Output tokens include hidden reasoning, generated code, and tool-call JSON, not only visible prose.
- Cache-read tokens are discounted, so total processed tokens materially overstate effective cost.
- The supplied receipt declares a one-hour cache-write TTL; the primary estimate follows it and preserves an all-five-minute sensitivity.
- No standardized local FPS/hardware receipt or blind visual-evaluation record is available.
Visible evidence gaps
- candidate prompt-delivery and one-shot protocol receipt
- standardized FPS, browser, viewport, and hardware receipt
- independently authored multi-cycle acceptance suite
- blind-evaluation record
Public result only
This result keeps its public summary and evidence, but it does not currently have a matching Builder test with prompts, projects, Harness workflows or skills.
