Same 1M tokens, very different cost — pick a plan by your context profile

30% cache hit, regular work hours — a good comparison baseline.

Fresh 50% · Cached 30% · Response 20%

Loading comparison data…