1. Count embedding records
Use the number of chunks or items sent to the embedding model, not merely the number of source files.
2. Estimate average token length
Measure a representative sample after chunking and cleaning.
3. Set refresh frequency
Enter one for initial indexing or a higher value for repeated full re-embedding.
4. Apply token and processing rates
Enter the current embedding price and any separate pipeline expense.
5. Review volume and unit cost
Check total tokens, model cost, and effective cost per record.