INF2 Accelerated Computing
⚡ GPU AcceleratedHardware accelerators (GPUs, FPGAs) for specialized compute.
Name:
infAWS Inferentia Acceleratorsinf = AWS Inferentia AcceleratorsPurpose-built AWS Inferentia silicon designed for ultra-low latency, high-throughput machine learning inference.22nd Generation2 = 2nd GenerationAWS 2nd generation hardware platform with updated CPU/system architecture.Common use cases
- Machine learning training/inference
- High-performance computing
- Graphics rendering
Head-to-Head Architecture Comparisons
View all comparisons →Instance types (4)
| Instance | vCPU | Memory | GPU Count | GPU Memory | Memory Type | Architecture | From (Linux, $/hr) |
|---|---|---|---|---|---|---|---|
| inf2.xlarge | 4 | 16 GB | 1x | 32.0 GB | Standard | x86_64 | $0.7582 |
| inf2.8xlarge | 32 | 128 GB | 1x | 32.0 GB | Standard | x86_64 | $1.9679 |
| inf2.24xlarge | 96 | 384 GB | 6x | 192.0 GB | Standard | x86_64 | $6.4906 |
| inf2.48xlarge | 192 | 768 GB | 12x | 384.0 GB | Standard | x86_64 | $12.9813 |