Measured harness ledgerPublic result
Claude Fable 5

Explorable space-flight game — Claude Fable 5 Max

Build a responsive browser space-flight game with flight controls, a coherent star-system environment, lighting, assets, and a playable game loop.

Max reasoningHeadline result
Workflow cost
$151.55
Wall-clock
1:05:19.2 wall-clock
Processed tokens
92.10M processed
Record state
partial_token_timing_and_source_ledger
Public summary

Claude Fable 5 Max partial_token_timing_and_source_ledger ledger: 1:05:19.2 wall-clock, 92.10M processed, and $151.55 API-equivalent estimate, not a subscription invoice.

Run identity and stack
  • Result ID: space-flight-game-fable-5-max
  • Technical model: claude-fable-5
  • Provider: Anthropic Claude Code
  • Stack: Anthropic Claude Code
  • Stack: Technical model/configuration: claude-fable-5
  • Stack: Three.js / Vite game
  • Stack: Generated spaceship assets
  • Stack: Harness v1 space-flight prompt
Cost basis
  • If every cache write used the official one-hour rate instead, the total would be $167.36.
  • Primary cost uses 5-minute cache writes.
  • All-1-hour cache-write alternative: $167.36.
  • No deduplication audit was supplied for this usage receipt; the token and cost figures retain the supplied basis.
Primary artifact integrity
  • Kind: interactive-threejs-space-flight-game-entry-point
  • Path: artifacts/space-flight-game-fable-5-max/source/src/main.js
  • SHA-256: 7fcb9aa3801a71e100a6cc64807308dc23ef4dbf454475c3bebd9c9a7790a94f
Recorded caveats
  • Wall-clock is end-to-end latency including idle and tool time, not model-only compute.
  • Output tokens include hidden reasoning, code, and tool-call JSON.
  • Cache reads are billed at a deep discount, so total processed tokens substantially overstate cost.
  • The supplied receipt assumes 5-minute cache writes. It does not separately identify cache-write TTLs.
  • No deduplication audit was supplied with this metrics receipt; the token and cost figures retain the supplied basis.
  • The build passed, but no local FPS measurement, browser/hardware environment, final capture, or blind evaluation was supplied.
Visible evidence gaps
  • browser and hardware environment
  • local FPS measurement receipt
  • final capture
  • blind-evaluation record
Builder test available

This result is part of a Builder test. Open it for the exact prompt and any released projects, RemakeBench Harness workflows and production skills. Public proof and known evidence gaps stay visible here.

  • Kimi K3 Launch 002 · v1
RemakeBenchResearch console