Measured harness ledgerPublic result
GPT-5.6 Luna

Interactive Three.js voxel map viewer — GPT-5.6 Luna xhigh

Build a Three.js voxel map viewer with selectable Jungle Temple, Oasis Outpost, and Shrine Village maps, camera reset behavior, and OrbitControls verification.

xhigh reasoningHeadline result
Workflow cost
$1.93
Wall-clock
15:00 wall-clock
Processed tokens
5.23M processed
Record state
partial_token_timing_and_source_ledger
Public summary

GPT-5.6 Luna xhigh partial_token_timing_and_source_ledger ledger: 15:00 wall-clock, 5.23M processed, and $1.93 API-equivalent estimate, not a subscription invoice.

Run identity and stack
  • Result ID: interactive-voxel-map-viewer-gpt-5.6-luna-xhigh
  • Technical model: gpt-5.6-luna
  • Provider: OpenAI Codex
  • Stack: OpenAI Codex
  • Stack: Technical model/configuration: gpt-5.6-luna
  • Stack: Three.js voxel viewer
  • Stack: OrbitControls
  • Stack: Harness v1 voxel-viewer prompt
Cost basis
  • Requests with more than 272,000 input tokens use long-context pricing; all 69 supplied calls were short-context.
  • Separately priced tools and non-token services are excluded.
Primary artifact integrity
  • Kind: interactive-threejs-voxel-map-viewer-component
  • Path: artifacts/interactive-voxel-map-viewer-gpt-5.6-luna-xhigh/source/app/page.tsx
  • SHA-256: ffa1730e6f850a64f7916798a9eae198ea6c7c589a37d5b77a63da89f97fd4e3
Recorded caveats
  • Wall-clock is end-to-end latency, not model-only compute; it includes tool time and idle gaps between user turns.
  • Output tokens include hidden reasoning, visible prose/code, and tool-call JSON.
  • Cached input is deeply discounted, so total processed tokens overstate cost.
  • This is an API-equivalent estimate rather than the actual charge for a subscription-backed Codex session.
  • The supplied record observed Priority service tier, so this estimate uses Priority rates and is not directly price-comparable with Standard-tier estimates.
  • Separately priced tools and non-token services are excluded.
  • Cache-creation tokens are zero because the Codex transcript schema does not expose a cache-write field.
  • Subagent logs are excluded to avoid double-counting inherited parent context.
Visible evidence gaps
  • browser and hardware environment
  • local FPS
  • final capture
  • blind-evaluation record
Public result only

This result keeps its public summary and evidence, but it does not currently have a matching Builder test with prompts, projects, Harness workflows or skills.

RemakeBenchResearch console