INF2 Accelerated Computing
⚡ GPU AcceleratedHardware accelerators (GPUs, FPGAs) for specialized compute.
Name:
infAWS Inferentia Acceleratorsinf = AWS Inferentia AcceleratorsPurpose-built AWS Inferentia silicon designed for ultra-low latency, high-throughput machine learning inference.22nd Generation2 = 2nd GenerationAWS 2nd generation hardware platform with updated CPU/system architecture.Common use cases
- Machine learning training/inference
- High-performance computing
- Graphics rendering
Instance types (4)
| Instance | vCPU | Memory | GPU Count | GPU Memory | Memory Type | Architecture | From (Linux, $/hr) |
|---|---|---|---|---|---|---|---|
| inf2.xlarge | 4 | 16.0 GiB | 1x | 32.0 GB | Standard | x86_64 | $0.7582 |
| inf2.8xlarge | 32 | 128.0 GiB | 1x | 32.0 GB | Standard | x86_64 | $1.9679 |
| inf2.24xlarge | 96 | 384.0 GiB | 6x | 192.0 GB | Standard | x86_64 | $6.4906 |
| inf2.48xlarge | 192 | 768.0 GiB | 12x | 384.0 GB | Standard | x86_64 | $12.9813 |