Collect a representative day
Group work into light, normal and heavy tasks. Record how many of each occur, the provider-reported token usage and whether retries were required. Include background automation and CI usage if it runs under the same budget.
Calculate a range
Multiply input and output tokens by the current model rates, then add all tasks. Create low and high cases for days with fewer or more agent loops. For subscriptions, use the provider’s current plan cost and usage rules instead of pretending every interaction is directly token billed.
Cost calculator
Enter your current rates and measured average. Nothing is uploaded.
Reduce the repeatable waste
Daily savings are most reliable when they come from repeated patterns such as duplicated project instructions, oversized logs or irrelevant tools. Varion can measure eligible reductions per request and aggregate the provider-bound difference without changing provider prices.
Measurement checklist
- Choose a representative completed task, not an artificial one-line prompt.
- Record the selected model, provider input, cached input, output, retries and final result.
- Change one optimization mechanism at a time so the cause remains visible.
- Verify required identifiers, tool calls, code changes or business fields.
- Keep passthrough available when the reduced request does not pass.
How Varion fits
Varion Token Engine is a gateway and testing platform for reducing eligible input-token waste across supported AI traffic. It reports original and provider-bound input, keeps provider charges separate, and does not claim that every request can be reduced. New verified users receive 100,000 processed input tokens and 50 local test runs.
Frequently asked questions
Should I include failed tasks?
Yes. They consume time and may consume tokens, so cost per successful task should include retries.
Can I use this for a team?
Yes. Add the representative daily sessions for each role and include shared automation.