#prompt caching

Amazon Bedrock prompt caching write and read pricing multipliers

Bedrock cache writes ate 85 percent of one AI bill

Prompt caching is sold as a saving. It can also be the single largest line on an inference invoice, and the failure is quiet, because a cache write succeeds whether or not anything ever reads it back.

Fri Aug 21 2026 · 6 min read · 4 views

AI

Comparison of AI API token prices against real cost per completed task

The cheapest AI API is not the cheapest to run

Every comparison chart ranks AI APIs by price per million tokens. That number does not tell you what a task costs. A model priced at $0.14 per million input tokens can finish a job for more money than

Sat Aug 08 2026 · 7 min read · 2 views

AI