1. Enter vector population
Use the number of vectors expected in the active index, including planned near-term growth if appropriate.
2. Describe vector storage
Enter dimensions and bytes per dimension after the selected precision or quantization.
3. Add index overhead
Allow for graph links, IDs, metadata, workspaces, allocator fragmentation, and serving runtime.
4. Enter GPU capacity and benchmark
Use usable memory and measured queries per second for the intended index and search parameters.
5. Compare constraints
The displayed count is driven by whichever is larger: resident memory or target throughput.