inf1.xlarge
⚡ 1x GPUPart of the INF1 Accelerated Computing family — Hardware accelerators (GPUs, FPGAs) for specialized compute.
Name:
infAWS Inferentia Acceleratorsinf = AWS Inferentia AcceleratorsPurpose-built AWS Inferentia silicon designed for ultra-low latency, high-throughput machine learning inference.11st Generation1 = 1st GenerationAWS 1st generation hardware platform with updated CPU/system architecture.xlargeExtra Large (1x)xlarge = Extra Large (1x)4 vCPUs • 8.0 GiB RAM • EBS only • Up to 25 Gigabit Network • 1x GPUPrice (Linux)
$0.2280/ hr
~$166.44/mo
Compute
4vCPUs
2 cores • x86_64
Memory
8.0GiB
2.0 GiB / vCPU
Value
$0.0570/ vCPU-hr
$0.0285/GiB-hr
Network
25Gbps
Baseline: 5 Gbps
Storage & EBS
593.75MB/s
Max: 20,000 IOPS
Technical Specifications & Limits
Hardware specs, container thresholds, and system limitsCompute & Processor
ProcessorIntel Xeon Platinum 8275CL (Cascade Lake)
Architecturex86_64 (x86_64)
vCPUs / Threads4 vCPUs
Physical Cores2
Clock Speed—
Intel AVX / AVX2No
Intel Turbo BoostNo
Bare MetalNo (Virtualized)
Performance & Benchmarks
CoreMark Score65,773
CoreMark / vCPU16,443.3 / core
FFmpeg Transcoding26 FPS
CoreMark per Dollar288,478 score / $
Cost per vCPU-hr$0.0570
Cost per GiB-hr$0.0285
Containers & Kubernetes
EKS Max Pods38 pods
Max ECS Tasks40 tasks
Max ENIs4
IPv4/v6 per ENI10
Total Addressable IPs40 total IPs
CNI Prefix DelegationSupported (High Density)
Networking & Bandwidth
Network PerformanceUp to 25 Gigabit
Baseline Bandwidth5 Gbps
Burst / Peak Bandwidth25 Gbps
Enhanced NetworkingNo
EFA SupportNo
IPv6 SupportNative Dual-Stack Supported
Storage & EBS Optimization
Storage ConfigurationEBS only
Local NVMe DiskNone (EBS only)
EBS Baseline Throughput148.75 MB/s
EBS Max Throughput593.75 MB/s
EBS Baseline IOPS4,000
EBS Max / Burst IOPS20,000
Dedicated EBS Throughput875 Mbps
GPU & Hardware Acceleration
GPU ModelNVIDIA Accelerator
Accelerator Count1x GPU(s)
Total GPU Memory—
Memory ArchitectureGDDR / HBM
CUDA Compute Capability—
Acceleration RuntimesCUDA, TensorRT, cuDNN, PyTorch, PyJAX
Platform & Architecture Details
Hypervisor / SystemAWS Nitro System (PCIe Offload)
Nitro EnclavesSupported (Isolated security enclaves)
Normalization FactorNA
Generation StatusCurrent Generation (Recommended)
Date Introduced2019-12-03
Regional Pricing Matrix
Live on-demand rates across all global AWS cloud regionsLowest Price Region
$0.2280/ hr
US East (N. Virginia) (us-east-1)Highest Price Region
$0.3770/ hr
South America (Sao Paulo) (sa-east-1)Global Fleet Average
$0.2774/ hr
Across all active regionsRegional Coverage
22regions
Available worldwide| Region Name | Region Code | On-Demand Price |
|---|---|---|
Africa (Cape Town) | af-south-1 | $0.3020 |
Asia Pacific (Hong Kong) | ap-east-1 | $0.3520 |
Asia Pacific (Mumbai) | ap-south-1 | $0.2400 |
Asia Pacific (Seoul) | ap-northeast-2 | $0.2810 |
Asia Pacific (Singapore) | ap-southeast-1 | $0.3080 |
Asia Pacific (Sydney) | ap-southeast-2 | $0.2850 |
Asia Pacific (Tokyo) | ap-northeast-1 | $0.3080 |
AWS GovCloud (US-East) | us-gov-east-1 | $0.2880 |
AWS GovCloud (US-West) | us-gov-west-1 | $0.2880 |
Canada (Central) | ca-central-1 | $0.2540 |
EU (Frankfurt) | eu-central-1 | $0.2850 |
EU (Ireland) | eu-west-1 | $0.2540 |
EU (London) | eu-west-2 | $0.2670 |
EU (Milan) | eu-south-1 | $0.2670 |
EU (Paris) | eu-west-3 | $0.2670 |
EU (Stockholm) | eu-north-1 | $0.2420 |
Middle East (Bahrain) | me-south-1 | $0.2800 |
South America (Sao Paulo)Highest | sa-east-1 | $0.3770 |
US East (N. Virginia)Lowest | us-east-1 | $0.2280 |
US East (Ohio)Lowest | us-east-2 | $0.2280 |
US West (N. California) | us-west-1 | $0.2740 |
US West (Oregon)Lowest | us-west-2 | $0.2280 |
No regions matched your search filter.