Legal Discovery Audit Sample Size Estimator

Determine a statistically based quality-control sample for a defined legal discovery population. It helps compliance and legal operations teams plan a review sample when they want a stated confidence level and margin of error rather than an arbitrary percentage of the population.

The model uses the standard proportion sample-size equation with a finite population correction. It estimates statistical precision for a simple random sample; it does not determine whether a particular audit methodology is legally sufficient. Stratification, risk-based sampling, clustered records, or very rare exceptions may call for a different design.

Calculator inputs

items
%
%
Result
Recommended statistical sample size
Calculated sample before rounding
Share of population
Expected exception rate
Confidence level

1. Enter the population
Provide the total number of eligible items in the audit population after applying the scope definition.

2. Choose a confidence level
Select 90%, 95%, or 99% depending on the level of statistical confidence required for the planning exercise.

3. Set the margin of error
Enter the desired plus-or-minus precision in percentage points. Smaller margins require larger samples.

4. Estimate the exception rate
Use an expected proportion when you have a reasonable basis. Use 50% as a conservative choice because it produces the largest sample for a given confidence and margin.

5. Review the recommended sample
The result is rounded up to a whole item and cannot exceed the population size.

6. Confirm the sampling method
Draw items randomly or use a more appropriate documented design if the population has meaningful subgroups or risk concentrations.

Initial sample n₀ = z² × p × (1 − p) ÷ e² Finite-population sample n = n₀ ÷ [1 + (n₀ − 1) ÷ N]

Where z is the selected confidence-level critical value, p is the expected exception proportion as a decimal, e is the desired margin of error as a decimal, and N is the population size. The calculator rounds the final result upward to avoid understating the modeled sample requirement.

What the result means

The result is the rounded-up number of items to sample under the selected simple-random-sample assumptions.

Statistical sample size alone does not establish legal or regulatory sufficiency; use the sampling design appropriate to the audit objective and population.

Given
Population: 50,000 items
Confidence level: 95% (z = 1.96)
Margin of error: 4%
Expected exception rate: 50%

Calculation
n₀ = 1.96² × 0.50 × 0.50 ÷ 0.0400² = 600.25
n = 600.25 ÷ [1 + (600.25 − 1) ÷ 50,000] = 593.14
Round up = 594 items

Result
A simple random sample of 594 items meets the selected statistical inputs under this model. The audit design still needs to reflect the purpose, risk profile, and structure of the population.

Why does 50% produce a large sample?

For a binary proportion, p = 50% maximizes p × (1 − p), which makes the sample requirement most conservative when the true exception rate is unknown.

Is a 95% confidence level always required?

No. The appropriate confidence level depends on the audit objective, risk tolerance, contractual requirements, and any governing standard. The calculator lets you compare common planning choices.

Can I audit fewer items than the result?

You can choose a different design, but the stated margin and confidence assumptions would no longer match this simple-random-sample calculation. Risk-based or stratified methods can be more efficient when properly designed.

What if the population is very small?

The finite population correction reduces the calculated sample as the sample becomes a substantial share of the population. The result is also capped at the total population size.

Does this sample size prove compliance?

No. Sample size addresses statistical precision for the selected model. Compliance conclusions also depend on scope, sampling execution, evidence quality, exception handling, and the substantive requirement being tested.