1. Enter Model A quality
Use a normalized score from the same evaluation used for Model B.
2. Add Model A cost
Enter average total model cost for one completed task.
3. Enter Model A latency
Use end-to-end response time under representative load.
4. Repeat for Model B
Keep units and measurement conditions identical.
5. Review value scores
The higher score indicates a stronger balance under this formula.
6. Inspect raw differences
Check whether the recommendation depends on a small or meaningful gap.
7. Run sensitivity scenarios
Change quality, cost, or latency assumptions to see when the choice changes.