1. Enter workload size
Use the number of agent tasks expected in the planning period.
2. Break down the context
Enter prompt, retrieved context, tool-message, and expected output tokens per task.
3. Allow for retries
Use the retry rate to represent failed runs, validation retries, or agent loops that repeat a model call.
4. Review the budget
Compare adjusted tokens per task with the total workload budget and output-token share.