This is why ForgeCode invests heavily in prompt caching.
In one workspace over the last 7 days on Opus 4.7
- 407M: input tokens
- 382.9M: cache-read tokens
- 98.1%: cache read ratio
- 22.9×: write amortization
Amortization = tokens read back per token written to cache.
So every 1 token cached was reused ~23 times.
At public API pricing, that’s ~$2,035 without caching vs ~$333 with 5-minute caching.
~$1.7K saved in input-token cost alone.
