Worked examplesSizing for 50,000 to 5 million pages a month
Four realistic volumes, sized for a mixed scanned document set running analyse, identify and classify, with enough headroom to clear a full monthly backlog inside a working week.
Sizing for 50,000 to 5 million pages a month| Criterion | 50K pages/month | 250K pages/month | 1M pages/month | 5M pages/month |
|---|
| GPU workers | 1× L40S 48 GBsingle node | 2× L40S 48 GBtwo nodes | 6× L40S 48 GBthree nodes | 24× L40S 48 GBeight to twelve nodes |
|---|
| Total vCPU | 16includes control plane | 48 | 144 | 576 |
|---|
| Total RAM | 128 GB | 256 GB | 768 GB | 3 TB |
|---|
| Local NVMe | 2 TBrender cache and model weights | 6 TB | 16 TB | 60 TB |
|---|
| Object storage, year one | 4 TB | 16 TB | 60 TBresults, cache and audit trail | 300 TB |
|---|
| Sustained throughput | ~1,800 pages/hr | ~3,600 pages/hr | ~10,800 pages/hr | ~43,200 pages/hr |
|---|
| Time to clear one month | ~28 hours | ~69 hours | ~93 hoursunder four days of continuous run | ~116 hours |
|---|
| Matching topology | RA-01 single node | RA-02 HA cluster | RA-02 or RA-04 | RA-02 with DR |
|---|
| Licence impact of volume | None | None | Nonesized by footprint, not pages | None |
|---|
GPU workers
- 50K pages/month
- 1× L40S 48 GBsingle node
- 250K pages/month
- 2× L40S 48 GBtwo nodes
- 1M pages/month
- 6× L40S 48 GBthree nodes
- 5M pages/month
- 24× L40S 48 GBeight to twelve nodes
Total vCPU
- 50K pages/month
- 16includes control plane
- 250K pages/month
- 48
- 1M pages/month
- 144
- 5M pages/month
- 576
Total RAM
- 50K pages/month
- 128 GB
- 250K pages/month
- 256 GB
- 1M pages/month
- 768 GB
- 5M pages/month
- 3 TB
Local NVMe
- 50K pages/month
- 2 TBrender cache and model weights
- 250K pages/month
- 6 TB
- 1M pages/month
- 16 TB
- 5M pages/month
- 60 TB
Object storage, year one
- 50K pages/month
- 4 TB
- 250K pages/month
- 16 TB
- 1M pages/month
- 60 TBresults, cache and audit trail
- 5M pages/month
- 300 TB
Sustained throughput
- 50K pages/month
- ~1,800 pages/hr
- 250K pages/month
- ~3,600 pages/hr
- 1M pages/month
- ~10,800 pages/hr
- 5M pages/month
- ~43,200 pages/hr
Time to clear one month
- 50K pages/month
- ~28 hours
- 250K pages/month
- ~69 hours
- 1M pages/month
- ~93 hoursunder four days of continuous run
- 5M pages/month
- ~116 hours
Matching topology
- 50K pages/month
- RA-01 single node
- 250K pages/month
- RA-02 HA cluster
- 1M pages/month
- RA-02 or RA-04
- 5M pages/month
- RA-02 with DR
Licence impact of volume
- 50K pages/month
- None
- 250K pages/month
- None
- 1M pages/month
- Nonesized by footprint, not pages
- 5M pages/month
- None
Sizing assumes a mixed scanned set at roughly 2.0 seconds of GPU time per page with the analyse, identify and classify capability set. Enabling Ask indexing across the same corpus adds about 50% to GPU time — plan an extra worker per five, or accept a longer bulk window. These are engineering starting points to be validated on your hardware.