1. Choose the planning period
Enter the tickets expected during the same period used for your budget or quota.
2. Estimate AI turns
Count each assistant response as one turn and use the average among automated tickets.
3. Measure customer input
Enter the average tokens in the customer message or condensed conversation state sent per turn.
4. Measure assistant output
Use the average generated reply length, including any structured fields returned by the model.
5. Allow for context overhead
Add system instructions, retrieved knowledge, summaries, tool messages, and expected retries as a percentage.
6. Use the total for capacity planning
Compare the final token requirement with account limits and cost-per-token assumptions.