The following table lists the recommended instance types required for AI Gateway configurations.
| AWS AMI | GCP | ESXi VM | Concurrent Requests | Minimum Disk Size | |
|---|---|---|---|---|---|
| small | c6a.4xlarge1 (16 cores/32GB) | c2d-highcpu-16 | 16 cores/32GB | CPU ML: 8 RPS GPU LLM: 20 RPS DLPaaS: 80 RPS | 300GB |
| medium | c6a.8xlarge1 (32 cores/64GB) | c2d-highcpu-32 | 32 cores/64GB | CPU ML: 15 RPS GPU LLM: 20 RPS DLPaaS: 80 RPS | 300GB |
Newer instance generations usually offer better performance. Customers can choose newer families, such as c6i, c7i, c7a, c8i, or c8a, for improved throughput.
Requirement: The AI Gateway VM must be deployed on an AVX2-capable host.
Warning: The AI Gateway deployment will fail if the host does not meet the minimum specifications listed above.
Warning: The AI Gateway deployment will fail if the host does not meet the minimum specifications listed above.
DLP Appliance Sizing
The following table lists the recommended instance types for cloud deployments and the corresponding size configurations for on-premises environments of the DLP On Demand appliance. For more information, see /en/dlpondemandconfig.
| AWS | GCP | Azure | VM | Concurrent Requests | Minimum Disk Size | |
|---|---|---|---|---|---|---|
| small | c5ad.4xlarge | n2-standard-16 | Standard-F16s_v2 | 16 cores/32GB | up to 7 requests/second for an average file size of 250KB | 351GB |
| medium | c5ad.8xlarge | n2-standard-32 | Standard-F32s_v2 | 32 cores/64GB | up to 30 requests/second for an average file size of 250KB | 351GB |
| Large | c5ad.16xlarge | n2-standard-64 | Standard-F64fs_v2 | 64 cores/128GB | up to 50 requests/second for an average file size of 250KB | 351GB |

