1. Enter cache behavior
Provide the expected hit rate and measured or benchmarked hit and miss latencies.
2. Set parallel capacity
Enter the number of requests that can be processed simultaneously.
3. Choose utilization
Use a utilization target below 100% to leave room for traffic variation.
4. Review latency and throughput
Compare the weighted latency with the estimated sustainable requests per second.