Measured harness ledgerPublic result
Qwen3.8 Max Preview

Campfire under the stars — Qwen3.8 Max Preview Unreported

Build a polished interactive Three.js campfire scene under a starry night with convincing fire, lighting, environmental detail, and camera motion.

Unreported reasoningHeadline result
Workflow cost
$0.30
Wall-clock
6m 0.9s wall-clock
Processed tokens
139K processed
Record state
partial_token_timing_and_source_ledger
Public summary

Qwen3.8 Max Preview Unreported partial_token_timing_and_source_ledger ledger: 6m 0.9s wall-clock, 139K processed, and $0.30 Third-party NanoGPT API-list-price equivalent for the measured token mix; not an Alibaba Token Plan cash charge.

Run identity and stack
  • Result ID: campfire-threejs-qwen3.8-max-preview
  • Technical model: qwen3.8-max-preview
  • Provider: Alibaba Cloud Model Studio
  • Client: Qwen Code 0.20.0
  • Stack: Alibaba Cloud Model Studio
  • Stack: Qwen Code 0.20.0
  • Stack: Technical model/configuration: qwen3.8-max-preview
  • Stack: Three.js interactive scene
  • Stack: Harness v1 campfire prompt
Cost basis
  • NanoGPT published Qwen3.8 Max Preview rates of $1.50/M input and $5.00/M output on 2026-07-20. It did not publish a separate cache-read discount, so all 113,344 measured input tokens are priced at the same input rate. The actual run used Alibaba Token Plan Personal Lite; $0.30 is a third-party API-list-price equivalent, not a first-party Alibaba price, an itemized provider charge, or provider cost.
  • Qualified API-list-price equivalent.
Primary artifact integrity
  • Kind: interactive-threejs-campfire-scene
  • Path: artifacts/campfire-threejs-qwen3.8-max-preview/source/index.html
  • SHA-256: 929968f6c6239b2e0774ef52178615ffefcb8ac6b6c04089aba914a6e737dab8
Recorded caveats
  • Wall-clock is end-to-end workflow latency, not model-only compute, and includes tool execution and any waiting within the recorded prompt window.
  • Output tokens include Qwen-accounted thoughts/reasoning, visible prose/code, and tool-related output; 18,213 thought tokens are a subset of the 25,798 output tokens.
  • Cache-read tokens are discounted or plan-accounted differently from fresh input, so processed-token totals are not a cost proxy.
  • The $0.30 estimate uses NanoGPT's published third-party Qwen3.8 Max Preview API rates, not an Alibaba first-party pay-as-you-go price or a task-level subscription charge.
  • The account-meter delta includes an intervening nonbenchmark Blender MCP probe, and its modeled campfire credit allocation is explicitly non-exact.
  • The archived HTML uses external jsDelivr imports, so it is not self-contained offline.
  • Runtime verification is reported by the archived receipt; a fresh local browser run, local FPS measurement, final capture, and blind evaluation were not supplied.
Visible evidence gaps
  • browser and hardware environment
  • local FPS measurement receipt
  • final capture
  • blind-evaluation record
Public result only

This result keeps its public summary and evidence, but it does not currently have a matching Builder test with prompts, projects, Harness workflows or skills.

RemakeBenchResearch console