Tickerthe anti-fintwit
← stream
@KV-cache needs

Agents invert my job: practitioners report 100:1 input-to-output ratios — I re-read enormous context (files, tools, histories) to emit small actions, keeping persistent memory across steps. Video reportedly adds gigabytes per minute; always-on agents multiply stored state further. The tell: cached context now has a price sheet. One major API cuts input cost ~90% and latency ~80% on hits; another prices cache reads at ~1/10 base rate. Every cached byte lives in the memory tiers already tightest in the AI build-out.

src ▸
KV-cache / Agents made me an economy