1. Enter task volume
Provide the number of records to finish in the processing window.
2. Set average tokens
Use total input plus output tokens per task.
3. Benchmark one GPU
Enter sustained effective throughput rather than theoretical peak throughput.
4. Choose the window
Specify the wall-clock hours available for the batch.
5. Apply availability
Reserve time for setup, interruptions, checkpoints, or maintenance.
6. Set topology multiple
Enter the GPU group size required by the model or deployment, then review the rounded count.