Provider notes

LLM provider differences

The same model is not the same product everywhere, and these notes are about the differences that change a bill or a cache. Claude on Anthropic’s own API, on AWS Bedrock and on Google Vertex differs in authentication, model IDs, price per endpoint class and where the prompt cache lives. Bedrock’s two ways to authenticate decide whether anything in the request path can add a cache breakpoint.

A per-token price depends on the platform, endpoint class and billing mode, not just the model name. And providers differ in what a proxy can do about prompt caching at all: create hits, raise the odds, or only watch. Each note says what to check before you compare prices or move traffic.

Start here

LLM token cost is a function, not a number

What a per-token price depends on; the other notes are cases of it.

All 4 posts

To price one call on each platform side by side, use the token cost calculator.