
Why LLM Routing Has to Price Prompt Caching
Ranking providers on list prices picks the wrong one for cache-heavy workloads: two deployments of the same model can differ by 2.6x on the invoice once cache reads are billed. Here is how LLM Gateway's routing now blends cached input prices into provider selection, the math behind the defaults, and how to tune them.







