Measured harness ledgerPublic result
Qwen3.8 Max Preview

Infinite cathedral corridor shader — Qwen3.8 Max Preview Unreported · Attempt 2

Generate an infinite stained-glass cathedral corridor shader with convincing depth, architectural repetition, light, and motion.

Unreported reasoningHeadline result
Workflow cost
$5.32
Wall-clock
25m 9.2s wall-clock
Processed tokens
3.41M processed
Record state
published_artifact_runtime_ledger
Public summary

Qwen3.8 Max Preview Unreported published_artifact_runtime_ledger ledger: 25m 9.2s wall-clock, 3.41M processed, and $5.32 Qualified third-party NanoGPT API-list-price scenario for the measured token mix; not an Alibaba Token Plan cash charge.

Run identity and stack
  • Result ID: infinite-cathedral-shadertoy-qwen3.8-max-preview-attempt-2-isolated-vision
  • Technical model: qwen3.8-max-preview
  • Provider: Alibaba Cloud Model Studio
  • Client: Qwen Code 0.20.0
  • Attempt: 2
  • Stack: Alibaba Cloud Model Studio
  • Stack: Qwen Code 0.20.0
  • Stack: Technical model/configuration: qwen3.8-max-preview
  • Stack: Attempt 2
  • Stack: ShaderToy fragment shader
  • Stack: Harness v1 infinite cathedral prompt
Cost basis
  • No first-party Qwen3.8 cache rate was established, so cached input uses the published third-party input rate. This is not Alibaba pricing, provider cost, or an itemized subscription charge.
  • Qualified API-list-price equivalent.
Primary artifact integrity
  • Kind: shadertoy-webgl-fragment-shader
  • Path: artifacts/infinite-cathedral-shadertoy-qwen3.8-max-preview-attempt-2-isolated-vision/source/shader.glsl
  • SHA-256: 7b10d4d80bcff819e1183873892fc04de45a441e5732deb9e95de70303158e1a
Recorded caveats
  • Wall-clock is end-to-end workflow latency, not model-only compute time.
  • Output tokens include Qwen-accounted thoughts/reasoning, visible prose/code, and tool-related output; 36,650 thought tokens are a subset of 60,232 output tokens.
  • The actual captures are 1920×993; 1920×1080 was the requested Chrome window size.
  • The independent 19.87 FPS sample is Apple M1 Pro / ANGLE Metal evidence, not a standardized cross-machine performance comparison.
  • Task-required initial and optimized local FPS receipts and an RTX Pro 6000 1080p60 final were not supplied.
  • The 28-Credit estimate (conservative range 25–32) is a reset-aware, after-only Token Plan capacity observation rather than a cash charge or invoice-grade measurement.
  • No formal RemakeBench quality score or blind-voting result has been recorded; the retained independent review is nonbinding provenance.
Visible evidence gaps
  • initial local FPS measurement receipt
  • optimized local FPS measurement receipt
  • RTX Pro 6000 1080p60 final capture
  • blind-evaluation record
Public result only

This result keeps its public summary and evidence, but it does not currently have a matching Builder test with prompts, projects, Harness workflows or skills.

RemakeBenchResearch console