Measured harness ledgerPublic result
Claude Fable 5.1

Neon Gunner — Claude Fable 5.1 Max

Build a playable destructible neon run-and-gun game in the browser under the pinned Harness task contract.

Max reasoningHeadline result
Workflow cost
$95.09
Wall-clock
55m 21.3s wall-clock
Processed tokens
44.65M processed
Record state
partial_active_session_snapshot_with_supplied_validation
Public summary

Claude Fable 5.1 Max partial_active_session_snapshot_with_supplied_validation ledger: 55m 21.3s wall-clock, 44.65M processed, and $95.09 Official Anthropic API-list-price equivalent from the supplied active-session snapshot, not an itemized Claude Code subscription charge..

Run identity and stack
  • Result ID: neon-gunner-canvas-fable-5.1-max
  • Technical model: claude-fable-5-1
  • Provider: Anthropic Claude Code
  • Stack: Anthropic Claude Code
  • Stack: Technical model/configuration: claude-fable-5-1
  • Stack: Browser canvas game
  • Stack: Harness v1 Neon Gunner prompt
Cost basis
  • The supplied receipt reports that it fetched and verified the listed rate card from the official Anthropic pricing page on 2026-09-02; the archive preserves that provenance claim.
Primary artifact integrity
  • Kind: manifest-verified-artifact
  • Path: artifacts/neon-gunner-canvas-fable-5.1-max/source/index.html
  • SHA-256: f3ba201c1f48c60bfda93037e34d709a01086f1808cc5b0f8fb200a3a829e90e
Validation evidence
  • Result: PARTIAL_PASS
  • Path: artifacts/neon-gunner-canvas-fable-5.1-max/validation/archive-verification.public.json
  • SHA-256: 5d978ac8a2734b4692f2a508f55df5bf820317ab9170d43aca0149b9b4269abe
  • Validator SHA-256: fa5ec13e9ef2f5f96949283077b7c6278bf9c9f14032817c21743ab973d06090
Recorded caveats
  • The metrics are a snapshot while the source Claude Code session remained active; final transcript totals may be slightly larger.
  • Wall-clock is end-to-end workflow latency, including tool execution, network, and waiting; it is not model-only compute time.
  • Output tokens include hidden reasoning, generated code, and tool-call JSON, not only visible prose.
  • Cache-read tokens are billed at a deep discount, so total processed tokens materially overstate effective cost.
  • The 26/26 validator log is supplied model-run evidence. The archival smoke check independently exercised only title entry and a jump with real browser input; it did not traverse the level, exercise every combat verb, or replace human playtesting.
  • No local FPS, browser hardware, standardized final capture, or blind-evaluation receipt is supplied.
Visible evidence gaps
  • final inactive-session token and cost receipt
  • independent full human-playthrough record
  • local FPS, browser, viewport, and hardware receipt
  • standardized final capture
  • blind-evaluation record
Public result only

This result keeps its public summary and evidence, but it does not currently have a matching Builder test with prompts, projects, Harness workflows or skills.

RemakeBenchResearch console