History
- Jul 30, 2026Compute usage is now metered by flavor.
How ready are your agents for a bad day? Get your scoreHow ready are your AI agents for a bad day? Get your security readiness score in 2 minutes.
Get my score →Tokens, compute, and storage — rolled up per agent, per team, per model.
Introduced Apr 12, 2026 · updated Jul 30, 2026
rolled up per agent · per team · per model
Tokens, compute, and storage roll up per agent, per team, per model — a fleet becomes chargeable to the teams that run it, not one line item nobody owns.
Tokens are counted the way the provider counts them, so the number in the report is the number on the invoice.
Cached input is counted separately — cache hit ratio is the leading indicator of what a fleet actually costs.
Thirty minutes, your infrastructure, your stack. Or skip the call — it is one Helm release onto a cluster you already run.