TCO Analysis Dashboard
Directionally accurate total cost of ownership comparison between on-premises infrastructure, cloud, and SaaS AI APIs.
GPU Configuration
Use Case GPU Profiles
This requires 1 GPU servers
Affects both hardware pricing and cloud instance selection
GPU Configuration Summary
On-Premises Configuration
Hardware Breakdown
Calculated: 8 total GPUs ÷ 8 GPUs per server = 1 servers
GPU server power
System power per server. Enter your server’s rating, or use the platform default.
Auto-selected Ethernet network
$45,000Sizes switches and connections for your deployment.
Single-homed access
Network details
One server uses its internal GPU interconnect; no external compute fabric is required. In-band and management connectivity are priced separately. Single-homed access; 48 endpoint ports per switch. Uplink and WAN/router ports are quote-dependent. Equipment costs assume new purchases. Shared-fabric allocation depends on the site.
Conventional HGX 400GbE compute fabric: 1 plane, 0 leaf and 0 spine switches. Includes converged in-band and out-of-band management; dedicated storage fabric is opt-in.
Edit unit prices
Hardware Maintenance
Personnel
Facilities Configuration
Platform operating model
A few choices set the inferred platform, security, and observability footprint.
Recommended topology: 3 control nodes. Production Kubernetes uses a three-node control plane; large clusters add a dedicated management node.
Included platform services
Add prices for the highlighted services to include them in totals.
Shared controls
| Service | On-premises | Cloud IaaS |
|---|---|---|
| Platform & OS | ||
| Commercial OS support | Add price· 160 cores | Add price· 160 cores |
| Kubernetes platform support | Add price· 3 nodes | Add price· 3 nodes |
| Security & AI governance | ||
| Isovalent Enterprise | Add price· 3 nodes | Add price· 3 nodes |
| Observability | ||
| Splunk Observability Cloud | $540/yr· 3 hosts | $540/yr· 3 hosts |
Cloud Configuration
Instance Configuration
Pricing Configuration
Choose your preferred pricing model
Enter your negotiated enterprise discount
AWS Enterprise Support fee percentage
Additional Storage
Personnel
Cloud Software Configuration
Calculating Costs
Updating summary...
Cloud Infrastructure
On-Premises Infrastructure
Comparison
—
Key Assumptions
Physical server counts follow the selected topology; cloud instance counts use only an exact GPU-to-SKU mapping or a manual quote.
Management software for compute and networking is included in the per unit cost.
EC2 On-Demand prices come from AWS’s live pricing sources and are cached for up to 24 hours.
SaaS Token API pricing (OpenAI, Anthropic, Gemini, Bedrock) is updated periodically based on published rates.
Personnel costs include benefits and overhead at 30% of base salary.
GPU server power follows installed GPU TDP plus platform host overhead. Rack capacity uses peak power; training energy uses scheduled activity with a 30% idle-GPU allowance. Shared servers and switches use a 1.2 kW planning allowance each. PUE applies to electricity.
On-premises infrastructure has 24/7/365 100% utilization.
Token API pricing uses the input/output token ratio from GenAI Sizing inputs (avgInputTokens/avgOutputTokens) and requires GenAI Sizing calculator results.
Effective cost per 1M output tokens accounts for both compute-bound prefill and memory-bound decode phases using: (requests/sec × output_tokens/request), which accurately reflects workloads with varying input/output ratios.
Networking includes converged in-band and out-of-band management switches. Compute fabric follows the selected topology policy; dedicated storage fabric is opt-in. WAN uplinks and cables require site-specific quotes.
Each GPU uses its exact AWS instance price; unavailable or reserved prices require a matching quote.
Roadmap
Launch v1.0!
2-4 GPU configurations per node support
Add cost per 1M token pricing
Token API SaaS pricing (OpenAI, Anthropic, Gemini, Bedrock)
Advanced networking cost modelling