Measured harness ledgerPublic result
Claude Fable 5

Explorable space-flight game — Claude Fable 5 Medium

Build a responsive browser space-flight game with flight controls, a coherent star-system environment, lighting, assets, and a playable game loop.

Medium reasoningHeadline result
Workflow cost
$22.85
Wall-clock
18:37.7 wall-clock
Processed tokens
13.96M processed
Record state
partial_token_timing_and_source_ledger
Public summary

Claude Fable 5 Medium partial_token_timing_and_source_ledger ledger: 18:37.7 wall-clock, 13.96M processed, and $22.85 Anthropic first-party Claude API, standard global pricing.

Run identity and stack
  • Result ID: space-flight-game-fable-5-medium
  • Technical model: claude-fable-5
  • Provider: Anthropic Claude Code
  • Stack: Anthropic Claude Code
  • Stack: Technical model/configuration: claude-fable-5
  • Stack: Three.js / Vite game
  • Stack: Generated spaceship assets
  • Stack: Harness v1 space-flight prompt
Cost basis
  • Primary cost uses 5-minute cache writes.
Primary artifact integrity
  • Kind: threejs-space-flight-game-entry-point
  • Path: artifacts/space-flight-game-fable-5-medium/source/src/main.js
  • SHA-256: 526fce0d17cc98662dd5a2d0c3d81f4ba363cdd72e1258b1b1f1d7aa55002404
Recorded caveats
  • Wall-clock is end-to-end latency including tool execution and waits, not model-only compute.
  • Output tokens include hidden reasoning, code, and tool-call JSON, not just visible text.
  • Cache-read is billed at a deep discount, so total processed tokens overstate effective cost.
  • The 5-minute cache-write rate is the supplied pricing basis; a one-hour cache write would increase the total.
  • This is an API-equivalent estimate rather than the actual charge for a subscription-backed Claude Code session.
  • Browser and hardware environment, local FPS, final capture, and blind-evaluation evidence have not been supplied.
Visible evidence gaps
  • browser and hardware environment
  • local FPS
  • final capture
  • blind-evaluation record
Builder test available

This result is part of Builder tests. Open them for the exact prompt and any released projects, RemakeBench Harness workflows and production skills. Public proof and known evidence gaps stay visible here.

  • Tech Review 001 · v1
  • Kimi K3 Launch 002 · v1
RemakeBenchResearch console