Skip to contentInside

Unit economics

The method follows Unit Economics of Infrastructure in the SRE Handbook: divide spend by a unit the work already counts. The units here are days, API calls, tokens, commits, deploy runs and eval runs. Claude Code figures are list-price values from Claude Code's own cost_usd; Showback explains how they relate to the bill.

Summary

UnitValueRangeSource
Claude Code, USD per daymean 143.35, median 94.307 full days, 2026-10-02 to 2026-10-08Loki
Claude Code, USD per API call0.1852026-10-01 15:21 to 2026-10-09 00:00 UTCLoki
Claude Code, USD per million tokens0.72sameLoki
Claude Code, USD per commit4.29same, 238 commitsgit log of ten repositories
Runner time per deploy push53.5 s, USD 0 billed51 runs, 2026-10-07 12:36 to 2026-10-09 10:01 UTCGitHub Actions API
Bedrock eval, USD per 1,000 calls0.37 to 1.84 by model and runruns of 2026-10-07eval records
Weekly decision-layer probeUSD 0.269 per runfirst run 2026-10-07probe record

Claude Code per day

Day (UTC)API callsUSDUSD per call
2026-10-01 (from 15:21)15617.560.113
2026-10-0213762.040.453
2026-10-031,396194.450.139
2026-10-0429786.410.291
2026-10-056353.410.848
2026-10-06175118.600.678
2026-10-072,754394.230.143
2026-10-0852794.300.179
Total5,5051,021.010.185

Cost per call varied by a factor of 7.5. The two days with the highest cost per call (2026-10-05 and 2026-10-06) also had the longest contexts: on average 411,000 and 368,000 cache-read tokens per call. The busiest day, 2026-10-07, averaged 182,000. A count of calls alone does not track spend.

Query, as on the dashboard (request_sum() in grafana/build.py), evaluated at the end of each UTC day:

sum(sum_over_time({service_name=~"claude-code.*"} | event_name="api_request" | keep cost_usd | unwrap cost_usd [1d]))

Claude Code per model

ModelAPI callsUSDShare of spendUSD per call
claude-opus-5-53,932506.5849.6%0.129
claude-fable-5-11,142442.8743.4%0.388
claude-opus-4-826466.116.5%0.250
claude-sonnet-5-5873.980.4%0.046
claude-haiku-4-5-20251001801.470.1%0.018

Range: 2026-10-01 15:21 to 2026-10-09 00:00 UTC. The same per-model spend is on the public Claude Code dashboard.

Tokens

Token classTokens
Input2,069,207
Output6,207,762
Cache read1,368,643,558
Cache write49,998,919
Total1,426,919,446

Cache reads were 96.3% of input-side tokens (input, cache read and cache write). On Opus 5.5 a cache read is priced at USD 0.20 per million tokens against USD 4.00 for uncached input (Claude API list prices, October 2026). The dashboard shows the same share as "cache read share".

Per commit

238 commits by the one author in the same range: 68 in the three public repositories (machinebehavior.io 50, tychat.io 10, agent-observability 8) and 170 in private or local repositories, 126 of them in the knowledge vault. USD 1,021.01 / 238 = USD 4.29 per commit.

A commit is a coarse unit. One commit can fix a typo or build 336 pages, and Claude Code spend also pays for research and drafts that reach no commit in the range. The figure is useful as a trend over several weeks; a single change has no fixed price.

Per deploy

The Conformity workflow runs the gate and the Pages deploy as two jobs (Deploy pipeline). Its 51 push runs from 2026-10-07 12:36 to 2026-10-09 10:01 UTC used 2,729 job-seconds: 53.5 s of runner time per push. The median time from push to finished run was 48 s. Billed: USD 0, as the repository is public and the run timing endpoint reports 0 billable ms. At the private-repository rate of USD 0.006 per minute, the same time would cost USD 0.27 in total, or USD 0.005 per push.

Evals on Bedrock

RunModelCallsUSDUSD per 1,000 calls
smoke-02qwen3-coder-30b3600.21210.59
full-01qwen3-coder-30b2,1601.30340.60
full-02, short stanceqwen3-coder-30b7200.29520.41
v5 smokegpt-oss-120b4800.26550.55
v5 smokeDeepSeek V3.24800.70871.48
v5 smokeKimi K2.54800.88181.84
v5 fullgpt-oss-120b2,8801.66740.58
v5 fullDeepSeek V3.22,8804.20591.46
v5 fullKimi K2.52,8805.13191.78
weekly probeqwen3-coder-30b7200.2690.37
Total14,04014.941.06

All runs on 2026-10-07. The harness sums the token counts of each reply at the prices it declares (AWS Pricing API, on-demand, 2026-10-07) and stops starting new conversations once the projected spend would cross the run's cap. Results: Experiment 04 and Weekly decision-layer probe.

The model calls under test cost USD 14.94. The Claude Code work that built, ran and published the experiment on the same day is inside that day's USD 394.23 and cannot be split out: the events carry no project label (Cost model).

Local LLM

1,716 requests with 1.30 million input and 66,883 output tokens, 2026-10-01 to 2026-10-09 00:00 UTC (ollama-exporter). The GPU used 2.22 kWh in the same range (mean 11.6 W). Its lowest reading was 8.98 W; at that draw for the whole range it would have used 1.72 kWh. That leaves about 0.50 kWh above the floor, or about 0.29 Wh per request. That figure is an estimate: it charges every watt above the minimum to the LLM. The money cost depends on the tariff, which is not published.

Benchmark

The Claude Code documentation gives an enterprise average of "around $13 per developer per active day", with 90% of users below USD 30 (Manage costs). This platform's mean is USD 143.35 per day at list price for one person. The comparison has limits: this person runs up to about 12 Claude Code processes at once (week to 2026-10-09), and 93% of the spend is on Opus 5.5 and Fable 5.1.