LLM cost measurement
What we found when we checked our own numbers against what actually happened. One audit compared every figure on Vigil’s dashboard with the code that computes it: nine definitions were wrong, including a savings total that was really a rolling window.
Another began with a single call that cost three cents more on the customer’s invoice than in our figures, and found two kinds of cost that token-based metering never sees: per-request fees such as web search, and tokens billed outside a response’s top-level usage. Each post says how the finding was checked, and what to look for in your own figures.
Start here
Nine AI dashboard metrics that were quietly wrong
The audit of every figure on our own dashboard.
All 2 posts
Anthropic web search pricing: the cost nobody meters
Anthropic bills web search at $10 per 1,000 requests, and advisor, fallback and compaction tokens outside top-level usage. Token-only metering misses both.
Nine AI dashboard metrics that were quietly wrong
We checked every figure on our AI monitoring dashboard against the code behind it. Nine definitions failed, one a savings total that was a rolling window.
To check your own figures against your invoice, the token cost calculator prices a call from the same rate table the proxy uses.