1. Set the record target
Enter the number of accepted synthetic records you want at the end of the run.
2. Estimate prompt size
Use the average token count for instructions, examples, metadata, and source context sent with each generation.
3. Estimate output size
Enter the expected tokens in one generated record before post-processing.
4. Include candidate generations
Specify how many candidate outputs are produced per accepted record.
5. Add a reserve
Allow for retries, safety-filter rejections, malformed output, or quality-screen failures.
6. Review the budget
Use total planned tokens for quota requests and cost modeling, then refine the assumptions after a pilot batch.