Analysis 01
How prompt caching actually changes the LLM cost math
Anthropic charges 10% of input for cache reads, with a 1.25x write fee. OpenAI auto-caches above 1024 tokens. The math changes which LLM is cheapest -- here is when.
13 posts tagged "prompt-caching".
Every post tagged "prompt caching" in the journal -- page 2 of 2.
Showing 1 of 13 posts — page 2 of 2
Anthropic charges 10% of input for cache reads, with a 1.25x write fee. OpenAI auto-caches above 1024 tokens. The math changes which LLM is cheapest -- here is when.