Measured harness ledgerPublic result
Claude Fable 5

High-voxel-density Jungle Temple diorama — Claude Fable 5 Medium

Create a complete inspectable Blender voxel diorama of a jungle temple at a fixed 0.05-unit voxel resolution through the Blender MCP workflow.

Medium reasoningHeadline result
Workflow cost
$10.10
Wall-clock
12:04.2 wall-clock
Processed tokens
2.52M processed
Record state
partial_token_timing_and_artifact_ledger
Public summary

Claude Fable 5 Medium partial_token_timing_and_artifact_ledger ledger: 12:04.2 wall-clock, 2.52M processed, and $10.10 Anthropic first-party Claude API, standard global pricing.

Run identity and stack
  • Result ID: jungle-temple-fable-5-medium
  • Technical model: claude-fable-5
  • Provider: Anthropic Claude Code
  • Stack: Anthropic Claude Code
  • Stack: Technical model/configuration: claude-fable-5
  • Stack: Blender MCP
  • Stack: 0.05-unit voxel grid
  • Stack: Inspectable .blend artifact
Cost basis
  • Primary cost uses 5-minute cache writes.
Primary artifact integrity
  • Kind: blender-scene
  • Path: artifacts/jungle-temple-fable-5-medium.blend
  • SHA-256: d9448b31f645dc0555697d65a7861d07832a55ca47aaa0d9c8071525bd98fcdf
Recorded caveats
  • Wall-clock is end-to-end latency including tool execution and user think-time, not model-only compute.
  • Output tokens include hidden reasoning, code, and tool-call JSON, not just visible text.
  • Cache-read tokens are billed at a deep discount, so total processed tokens greatly overstate effective cost.
  • The primary total uses the supplied five-minute cache-write rate; a one-hour cache write would increase the cost.
  • Blender and render environment, final capture metadata, and blind-evaluation evidence have not been supplied.
Visible evidence gaps
  • Blender and render environment
  • final capture metadata
  • blind-evaluation record
Builder test available

This result is part of Builder tests. Open them for the exact prompt and any released projects, RemakeBench Harness workflows and production skills. Public proof and known evidence gaps stay visible here.

  • Tech Review 001 · v1
  • Kimi K3 Launch 002 · v1
RemakeBenchResearch console