$43.97 Sol API-equivalent estimate versus $107.84 Fable combined estimate (API-equivalent estimate + estimated API cost). Fable cost 2.45× as much.
GPT-5.6 vs Claude Fable 5
Same frozen task prompt. One request per model. Seven workflows. Every recorded dollar.
Same target. Frozen prompt. Disclosed stack. Honest result.
- Lower total cost
- GPT-5.6 Sol Ultra$43.97 API-equivalent estimate
- Faster completion
- Claude Fable 5 Medium97:00.9 total workflow time
No synthetic overall winner. Quality and community preference remain separate evidence categories.
- 7benchmark prompts
- 28recorded runs
- 4configurations
- Disclosedestimated API costs
- 5 activeof seven blind matchups
Audit the workflow behind this result.
Get the 7 exact frozen prompts, 28 available ledgers, four configurations, and cost/time exports from Tech Review 001.
Verdict first. Receipts immediately after.
Cost and time are earned from seven public run ledgers per headline configuration. Missing quality, preference, and repeat evidence stays missing.
97:00.9 versus 185:12.7. Sol took 1.91× as long.
Fable produced some of the strongest high-end environments in the suite, especially Storm City and Space Flight. Sol Ultra won Cathedral and often came close for $43.97 versus Fable’s $107.84, while taking 1.91× as long. With one recorded run per configuration per test, reliability is not established.
No percentage appears before a real minimum sample exists.
Repeat-run evidence has not been attached.
One request. No rescue pass.
- One user request per run, with no human follow-up or manual correction after submission.
- The benchmark target and frozen task prompt matched within each round.
- Provider-specific agent stacks could use their disclosed tools and internal workflows; the stacks were not identical.
- Costs are estimates based on recorded token usage and each record’s cited pricing basis—not subscription invoices.
- Published benchmark versions stay immutable, and missing evidence remains visible.
Sponsors can support RemakeBench. They cannot buy benchmark outcomes, rankings, blind votes, or editorial verdicts.
Open the public methodology and evidence →Every round, every recorded configuration.
Fable Medium and Sol Ultra lead the comparison. Terra Ultra and Luna extra-high remain visible as capability and value references.
Infinite Cathedral Corridor
Create an infinite stained-glass cathedral corridor shader with convincing depth, architectural repetition, light, and motion.
Sol Ultra is the quality winner. Terra is the value surprise, while Luna shows the capability floor.


Post-hoc RTX PRO 6000 evidence verified
1920×1080 · 60 fps · 900 frames
Source hash matched · hardware and renderer attested
Deterministic offline capture + separate sustained-performance test
- Uncapped trials
- 1512.39 · 1510.32 · 1507.91 FPS
- Frame time
- p50 0.656 ms · p95 0.731 ms · p99 0.755 ms
- Trial variability
- 0.148% sample CV
- Sustained 60 fps
- Pass · 0 late frames · 0 deadline misses
- Classification
qualifying_posthoc_reproduction
Three post-hoc performance trials of one generated artifact; model-generation sample n=1.
- Uncapped trials
- 387.75 · 387.62 · 387.46 FPS
- Frame time
- p50 2.586 ms · p95 2.631 ms · p99 2.651 ms
- Trial variability
- 0.037% sample CV
- Sustained 60 fps
- Pass · 0 late frames · 0 deadline misses
- Classification
qualifying_posthoc_reproduction
Three post-hoc performance trials of one generated artifact; model-generation sample n=1.
Fable produced 3.90× Sol’s measured post-hoc EGL throughput on this task.The raw FPS number is not a visual-quality verdict, and these two opposing task results do not establish either model as generally more performant.
Post-hoc uncapped headless EGL/OpenGL throughput on an NVIDIA RTX PRO 6000 at 1920×1080. This is not the original local FPS or native TWIGL/ShaderToy browser performance.
Fable Medium
- Workflow cost
- $19.17 estimated API cost
- Workflow time
- 15:38.1 wall-clock
- Processed tokens
- 6.95M
- Reasoning
- Medium
Anthropic Claude Code · ShaderToy fragment shader · Harness v1 infinite cathedral prompt
User-supplied report using official Anthropic API rates; 5-minute cache-write assumption
Sol Ultra
- Workflow cost
- $5.61 API-equivalent estimate
- Workflow time
- 38:22.4 wall-clock
- Processed tokens
- 5.78M
- Reasoning
- Ultra
OpenAI Codex · ShaderToy fragment shader · Harness v1 infinite cathedral prompt
API-equivalent estimate, not a subscription invoice
Same target, disclosed alternative configurations.
Neo-Gothic Storm City
Create a storm-lashed neo-gothic city shader with legible architecture, atmosphere, lighting, water, and motion.
Fable wins this round. Its ocean, fog and lightning create the more convincing storm, even if some towers fracture under closer inspection.


Post-hoc RTX PRO 6000 evidence verified
1920×1080 · 60 fps · 900 frames
Source hash matched · hardware and renderer attested
Deterministic offline capture + separate sustained-performance test
- Uncapped trials
- 822.00 · 821.85 · 821.41 FPS
- Frame time
- p50 1.217 ms · p95 1.322 ms · p99 1.379 ms
- Trial variability
- 0.037% sample CV
- Sustained 60 fps
- Pass · 0 late frames · 0 deadline misses
- Classification
qualifying_posthoc_reproduction
Three post-hoc performance trials of one generated artifact; model-generation sample n=1.
- Uncapped trials
- 2049.47 · 2045.13 · 2041.02 FPS
- Frame time
- p50 0.491 ms · p95 0.546 ms · p99 0.565 ms
- Trial variability
- 0.207% sample CV
- Sustained 60 fps
- Pass · 0 late frames · 0 deadline misses
- Classification
qualifying_posthoc_reproduction
Three post-hoc performance trials of one generated artifact; model-generation sample n=1.
Sol produced 2.49× Fable’s measured post-hoc EGL throughput on this task.The raw FPS number is not a visual-quality verdict, and these two opposing task results do not establish either model as generally more performant.
Post-hoc uncapped headless EGL/OpenGL throughput on an NVIDIA RTX PRO 6000 at 1920×1080. This is not the original local FPS or native TWIGL/ShaderToy browser performance.
Fable Medium
- Workflow cost
- $9.09 estimated API cost
- Workflow time
- 8:15 wall-clock
- Processed tokens
- 1.93M
- Reasoning
- Medium
Anthropic Claude Code · TWIGL fragment shader · Harness v1 neo-gothic storm city prompt
Anthropic first-party Claude API, standard global pricing; 5-minute cache-write assumption
Sol Ultra
- Workflow cost
- $2.84 API-equivalent estimate
- Workflow time
- 22:10.9 wall-clock
- Processed tokens
- 2.81M
- Reasoning
- Ultra
OpenAI Codex · TWIGL classic fragment shader · Harness v1 neo-gothic storm city prompt
API-equivalent estimate, not a subscription invoice
Same target, disclosed alternative configurations.
Space Flight Game
Build a playable browser space-flight game with responsive controls, a coherent environment, lighting, assets, and a game loop.
Fable creates the more expansive environment and stronger visual detail. Terra has the best flight controls of the four.


Fable Medium
- Workflow cost
- $22.85 API-equivalent estimate
- Workflow time
- 18:37.7 wall-clock
- Processed tokens
- 13.96M
- Reasoning
- Medium
Anthropic Claude Code · Three.js source project · Generated GLB spaceship asset · Harness v1 space-flight prompt
Anthropic first-party Claude API, standard global pricing; 5-minute cache-write assumption
Sol Ultra
- Workflow cost
- $17.94 API-equivalent estimate
- Workflow time
- 37:44.2 wall-clock
- Processed tokens
- 25.9M
- Reasoning
- Ultra
OpenAI Codex · Three.js / app-router source bundle · Generated GLB spaceship asset · Harness v1 space-flight prompt
User-supplied API-equivalent pre-request snapshot
Same target, disclosed alternative configurations.
Campfire Under a Starry Night
Build a polished interactive Three.js campfire with convincing fire, light, environmental detail, and controllable camera motion.
Sol cost more than Fable. Luna was cheapest and included more controls; no clean quality winner was claimed.


Fable Medium
- Workflow cost
- $5.12 estimated API cost
- Workflow time
- 4:24.6 wall-clock
- Processed tokens
- 1.14M
- Reasoning
- Medium
Anthropic Claude Code · Three.js interactive scene · Harness v1 campfire prompt
Anthropic first-party Claude API, standard global pricing; 5-minute cache-write assumption
Sol Ultra
- Workflow cost
- $6.86 API-equivalent estimate
- Workflow time
- 23:08.4 wall-clock
- Processed tokens
- 8.72M
- Reasoning
- Ultra
OpenAI Codex · Three.js source project · Harness v1 campfire prompt
API-equivalent estimate, not a subscription invoice
Same target, disclosed alternative configurations.
Jungle Temple
Create an inspectable high-density voxel jungle-temple diorama in Blender at the requested fixed grid scale.
Fable creates the more complete and coherent scene. Sol has the preferred bridge design, while Terra and Luna expose the capability gap.


Fable Medium
- Workflow cost
- $10.10 estimated API cost
- Workflow time
- 12:04.2 wall-clock
- Processed tokens
- 2.52M
- Reasoning
- Medium
Anthropic Claude Code · Blender MCP · 0.05-unit voxel grid · Inspectable .blend artifact · Harness v1 jungle temple prompt
Anthropic first-party Claude API, standard global pricing; 5-minute cache-write assumption
Sol Ultra
- Workflow cost
- $3.23 API-equivalent estimate
- Workflow time
- 26:39.7 wall-clock
- Processed tokens
- 2.56M
- Reasoning
- Ultra
OpenAI Codex · Blender MCP · 0.05-unit voxel grid · Inspectable .blend artifact · Harness v1 jungle temple prompt
API-equivalent estimate, not a subscription invoice
Same target, disclosed alternative configurations.
Shrine Village
Create an inspectable voxel shrine-village scene in Blender with coherent composition, architecture, landscaping, and detail.
Sol overperforms on value at $1.98 versus Fable’s $20.28, but floating roof artifacts and an unusable bridge prevent a clean quality win.


Fable Medium
- Workflow cost
- $20.28 estimated API cost
- Workflow time
- 18:37.5 wall-clock
- Processed tokens
- 5.94M
- Reasoning
- Medium
Anthropic Claude Code · Blender MCP · 0.05-unit voxel grid · Inspectable .blend artifact · Harness v1 shrine village prompt
Anthropic first-party Claude API, standard global pricing; 5-minute cache-write assumption
Sol Ultra
- Workflow cost
- $1.98 API-equivalent estimate
- Workflow time
- 20:36.2 wall-clock
- Processed tokens
- 1.53M
- Reasoning
- Ultra
OpenAI Codex · Blender MCP · 0.05-unit voxel grid · Inspectable .blend artifact · Harness v1 shrine village prompt
API-equivalent estimate, not a subscription invoice
Same target, disclosed alternative configurations.
Oasis Outpost
Create an inspectable voxel oasis outpost in Blender with a readable settlement, terrain, vegetation, water, and environmental detail.
Fable creates the broadest environment. Terra preserves a complete, readable oasis at $1.33; no clean quality winner was claimed.


Fable Medium
- Workflow cost
- $21.23 estimated API cost
- Workflow time
- 19:23.7 wall-clock
- Processed tokens
- 6.57M
- Reasoning
- Medium
Anthropic Claude Code · Blender MCP · 0.05-unit voxel grid · Inspectable .blend artifact · Harness v1 oasis outpost prompt
Anthropic first-party Claude API, standard global pricing; 5-minute cache-write assumption
Sol Ultra
- Workflow cost
- $5.51 API-equivalent estimate
- Workflow time
- 16:30.9 wall-clock
- Processed tokens
- 1.85M
- Reasoning
- Ultra
OpenAI Codex · Blender MCP · 0.05-unit voxel grid · Inspectable .blend artifact · Harness v1 oasis outpost prompt
Priority API short-context rates as an API-equivalent scenario because the supplied transcript observed service_tier=priority; this is not proof of an API invoice.
Same target, disclosed alternative configurations.
Judge the outputs before you see the model.
Explore each real build for at least 15 seconds. Sign in only when you record the vote. Identity, cost, tokens, time, and evidence reveal afterward.
Take the exact prompts, available ledgers, configurations, and exports with you.
Get prompts + ledgers