Anthropic through the Vigil proxy
Anthropic’s own API for the Claude models, at api.anthropic.com, with explicit prompt caching through cache_control breakpoints. Point your SDK's base URL at Vigil, add one header, and every Anthropic call is logged with its cost, tokens, latency, errors and agent, then forwarded to https://api.anthropic.com unchanged, unless you switch prompt-cache optimisation on.
What Vigil records per Anthropic call
One row per call: the model id as Anthropic returned it, input and output tokens, cache reads and writes where the response reports them, latency, the HTTP status and the provider's error code when it fails, and the agent named in the X-Vigil-Agent header. Likely prompt injections and personal data in prompts are flagged on the Errors page. A cost spike is flagged when a call costs over three times the agent's seven-day median.
Cost is computed from the proxy's rate registry, which prices 13 models on this platform today. A model the registry has no row for is logged with its cost left empty, never estimated.
- claude-3-5-haiku
- claude-fable-5
- claude-fable-5-1
- claude-haiku-4-5-20251001
- claude-opus-4-5
- claude-opus-4-6
- claude-opus-4-7
- claude-opus-4-8
- claude-opus-5
- claude-sonnet-4
- claude-sonnet-4-5
- claude-sonnet-4-6
- claude-sonnet-5
Price per token, by model: Claude Opus 5, Claude Sonnet 5, Claude Haiku 4.5, Claude Fable 5.1, Claude Opus 4.8, Claude Sonnet 4.6.
Setup
Every route has the shape https://api.vigil.tools/{user_id}/{provider}/{path}. For Anthropic the provider segment is anthropic, so the base URL is https://api.vigil.tools/{user_id}/anthropic and a request path such as /v1/messages is forwarded unchanged.
Two headers: X-Vigil-Key, your Vigil key from the dashboard, and X-Vigil-Agent, the name the dashboard groups this traffic under. Your Anthropic key stays where it was, in ANTHROPIC_API_KEY.
import Anthropic from "@anthropic-ai/sdk";
const client = new Anthropic({
apiKey: process.env.ANTHROPIC_API_KEY, // unchanged
baseURL: "https://api.vigil.tools/{user_id}/anthropic",
defaultHeaders: {
// Shown once at creation. Treat it like any other secret.
"X-Vigil-Key": "vk_your_vigil_key",
"X-Vigil-Agent": "my-agent",
},
});
const message = await client.messages.create({
model: "claude-haiku-4-5",
max_tokens: 1024,
messages: [{ role: "user", content: "Hello!" }],
});
// Using the Vercel AI SDK instead? Same two changes:
// import { createAnthropic } from "@ai-sdk/anthropic";
// const anthropic = createAnthropic({ baseURL: "https://api.vigil.tools/{user_id}/anthropic", headers: { "X-Vigil-Key": "vk_your_vigil_key" } });import anthropic
client = anthropic.Anthropic(
base_url="https://api.vigil.tools/{user_id}/anthropic",
default_headers={
# Shown once at creation. Treat it like any other secret.
"X-Vigil-Key": "vk_your_vigil_key",
"X-Vigil-Agent": "my-agent",
},
)
message = client.messages.create(
model="claude-haiku-4-5",
max_tokens=1024,
messages=[{"role": "user", "content": "Hello!"}],
)
# Using LangChain instead? Same two changes:
# from langchain_anthropic import ChatAnthropic
# llm = ChatAnthropic(model="claude-haiku-4-5", anthropic_api_url="https://api.vigil.tools/{user_id}/anthropic",
# default_headers={"X-Vigil-Key": "vk_your_vigil_key"})curl https://api.vigil.tools/{user_id}/anthropic/v1/messages \
-H "content-type: application/json" \
-H "x-api-key: $ANTHROPIC_API_KEY" \
-H "anthropic-version: 2023-06-01" \
-H "X-Vigil-Key: vk_your_vigil_key" \
-H "X-Vigil-Agent: my-agent" \
-d '{
"model": "claude-haiku-4-5",
"max_tokens": 1024,
"messages": [{ "role": "user", "content": "Hello!" }]
}'Replace {user_id} and vk_your_vigil_key with the values on your Connect page. The snippets above are generated by the same code as that page.
Prompt caching on Anthropic
Vigil adds cache breakpoints to your requests — it reads the prompt structure and marks where the reusable prefix ends, so the provider caches exactly that much. This is the mechanism the Optimise page measures and reports savings for.
With optimisation On: Vigil adds cache markers to the stable start of the prompt. In Shadow, Vigil measures what caching would have saved and changes nothing. Off records the call and nothing more.
Questions
- Does Vigil see or store my Anthropic API key?
- No, not the key itself. Your Anthropic key passes through the proxy with the request and is never written to disk. A call record can keep a short one-way digest of the credential, so that one key’s cache is kept apart from another’s; it cannot be turned back into the key. Vigil identifies you by the X-Vigil-Key header and your user id in the URL.
- Does routing Anthropic through Vigil add latency?
- Some. The proxy runs on Cloudflare’s edge, so the added hop is short, but reading the request body to place cache markers takes time, and there is a bounded lookup for your account state. The dashboard shows total latency per call. For long calls, stream: Cloudflare’s edge closes non-streaming connections after roughly 125 seconds.
- Which Anthropic models does Vigil price?
- 13 models with a published rate in the registry: claude-3-5-haiku, claude-fable-5, claude-fable-5-1, claude-haiku-4-5-20251001, claude-opus-4-5, claude-opus-4-6, claude-opus-4-7, claude-opus-4-8, claude-opus-5, claude-sonnet-4, claude-sonnet-4-5, claude-sonnet-4-6, claude-sonnet-5. A call to any other Anthropic model is logged and monitored, with its cost left empty rather than guessed.
Other providers: AWS Bedrock · Google Vertex · OpenAI · Google Gemini · Mistral · xAI Grok · DeepSeek · Together AI · Fireworks AI · Groq · Cerebras · Baseten · Moonshot · Z.ai · Cloudflare Workers AI · OpenRouter