Measured harness ledgerPublic result
GPT-5.6 SolOff-road Mud Game — GPT-5.6 Sol Max
Build a one-request procedural off-road driving game with vehicle physics, independent suspension, streamed terrain, mud feedback, and multiple cameras.
Max reasoningHeadline result
- Workflow cost
- $19.10
- Wall-clock
- 1h 2m 40.9s wall-clock
- Processed tokens
- 25.94M processed
- Record state
- partial_token_timing_source_build_ledger
Public summary
GPT-5.6 Sol Max partial_token_timing_source_build_ledger ledger: 1h 2m 40.9s wall-clock, 25.94M processed, and $19.10 Token-only standard API-equivalent estimate, not the actual Codex subscription charge.
Run identity and stack
- Result ID: off-road-driving-game-threejs-gpt-5.6-sol-max
- Technical model: gpt-5.6-sol
- Provider: OpenAI Codex
- Stack: OpenAI Codex
- Stack: Technical model/configuration: gpt-5.6-sol
- Stack: Three.js vehicle workflow
- Stack: Harness v1 off-road prompt
Cost basis
- The source receipt applies long-context rates only when a call exceeds 272,000 input tokens. It reports zero such calls.
- No separately priced tool or non-token cost is included because verifiable transcript usage counts were not supplied.
Primary artifact integrity
- Kind: ledger-checksummed-source-file
- Path: artifacts/off-road-driving-game-threejs-gpt-5.6-sol-max/source/app/ExpeditionGame.tsx
- SHA-256: 0b8d432e0de43e223ed85bfa1979f449063145567b22ab9fcd5f262379ec149e
Recorded caveats
- Wall-clock is end-to-end workflow latency, including tools and idle gaps; it is not model-only compute time.
- Output tokens include hidden reasoning, visible prose/code, and tool-call JSON.
- Cache-read tokens are discounted, so total processed tokens substantially overstate cost.
- The generated production output is deliberately not archived because its server-side build files contain a build-time prerender secret. The complete committed source required to regenerate it is archived instead.
- The production build passes, but the supplied npm test command fails two stale starter-scaffold assertions. No independent browser acceptance, sustained-drive, local-FPS, seam, or suspension-articulation receipt is archived.
- No final gameplay capture or blind evaluation is supplied.
Visible evidence gaps
- independent browser acceptance run
- local FPS, browser, viewport, and hardware receipt
- multi-minute sustained-drive and generation-seam receipt
- four-corner suspension-articulation receipt
- final gameplay capture
- blind-evaluation record
Public result only
This result keeps its public summary and evidence, but it does not currently have a matching Builder test with prompts, projects, Harness workflows or skills.
