Measured harness ledgerPublic result
Qwen3.8 Max PreviewInfinite cathedral corridor shader — Qwen3.8 Max Preview Unreported · Attempt 2
Generate an infinite stained-glass cathedral corridor shader with convincing depth, architectural repetition, light, and motion.
Unreported reasoningHeadline result
- Workflow cost
- $5.32
- Wall-clock
- 25m 9.2s wall-clock
- Processed tokens
- 3.41M processed
- Record state
- published_artifact_runtime_ledger
Public summary
Qwen3.8 Max Preview Unreported published_artifact_runtime_ledger ledger: 25m 9.2s wall-clock, 3.41M processed, and $5.32 Qualified third-party NanoGPT API-list-price scenario for the measured token mix; not an Alibaba Token Plan cash charge.
Run identity and stack
- Result ID: infinite-cathedral-shadertoy-qwen3.8-max-preview-attempt-2-isolated-vision
- Technical model: qwen3.8-max-preview
- Provider: Alibaba Cloud Model Studio
- Client: Qwen Code 0.20.0
- Attempt: 2
- Stack: Alibaba Cloud Model Studio
- Stack: Qwen Code 0.20.0
- Stack: Technical model/configuration: qwen3.8-max-preview
- Stack: Attempt 2
- Stack: ShaderToy fragment shader
- Stack: Harness v1 infinite cathedral prompt
Cost basis
- No first-party Qwen3.8 cache rate was established, so cached input uses the published third-party input rate. This is not Alibaba pricing, provider cost, or an itemized subscription charge.
- Qualified API-list-price equivalent.
Primary artifact integrity
- Kind: shadertoy-webgl-fragment-shader
- Path: artifacts/infinite-cathedral-shadertoy-qwen3.8-max-preview-attempt-2-isolated-vision/source/shader.glsl
- SHA-256: 7b10d4d80bcff819e1183873892fc04de45a441e5732deb9e95de70303158e1a
Recorded caveats
- Wall-clock is end-to-end workflow latency, not model-only compute time.
- Output tokens include Qwen-accounted thoughts/reasoning, visible prose/code, and tool-related output; 36,650 thought tokens are a subset of 60,232 output tokens.
- The actual captures are 1920×993; 1920×1080 was the requested Chrome window size.
- The independent 19.87 FPS sample is Apple M1 Pro / ANGLE Metal evidence, not a standardized cross-machine performance comparison.
- Task-required initial and optimized local FPS receipts and an RTX Pro 6000 1080p60 final were not supplied.
- The 28-Credit estimate (conservative range 25–32) is a reset-aware, after-only Token Plan capacity observation rather than a cash charge or invoice-grade measurement.
- No formal RemakeBench quality score or blind-voting result has been recorded; the retained independent review is nonbinding provenance.
Visible evidence gaps
- initial local FPS measurement receipt
- optimized local FPS measurement receipt
- RTX Pro 6000 1080p60 final capture
- blind-evaluation record
Public result only
This result keeps its public summary and evidence, but it does not currently have a matching Builder test with prompts, projects, Harness workflows or skills.
