TCO Analysis Dashboard
Directionally accurate total cost of ownership comparison between on-premises infrastructure, cloud, and SaaS AI APIs.
TCO Comparison: Cloud vs. On-Premises
Cloud Pricing Model
Analysis Period
Calculating TCO Analysis
Fetching cloud pricing data and computing cost comparisons for your configuration...
On-Premises Total Cost
$70,210.86/month avg.
Cloud Total Cost
$6,839.64/month avg.
Recommended Option
CLOUDCloud infrastructure is more cost-effective for your use case, saving approximately $2,281,364.00 (90.3%) over 3 years.
On-Premises Cost Breakdown
Cloud Cost Breakdown
GPU Configuration
Use Case GPU Profiles
This requires 1 GPU servers
Affects both hardware pricing and cloud instance selection
GPU Configuration Summary
On-Premises Configuration
Hardware Breakdown
Calculated: 8 total GPUs ÷ 8 GPUs per server = 1 servers
Auto-selected Ethernet network
$1,226,000Conventional HGX 400GbE compute fabric: 1 plane, 2 leaf and 2 spine switches. Includes converged in-band and out-of-band management; dedicated storage fabric is opt-in.
Adjust quote rates
Hardware Maintenance
Personnel
Facilities Configuration
Platform operating model
A few choices set the inferred platform, security, and observability footprint.
Recommended topology: 3 control nodes. Production Kubernetes uses a three-node control plane; large clusters add a dedicated management node.
Included platform services
Quantities are inferred. Quote-required items are visible but excluded from totals until priced.
Shared controls
| Service | On-premises | Cloud IaaS |
|---|---|---|
| Platform & OS | ||
| Commercial OS support | Quote required· 160 cores | Quote required· 160 cores |
| Kubernetes platform support | Quote required· 3 nodes | Quote required· 3 nodes |
| Security & AI governance | ||
| Isovalent Enterprise | Quote required· 3 nodes | Quote required· 3 nodes |
| Observability | ||
| Splunk Observability Cloud | $540/yr· 3 hosts | $540/yr· 3 hosts |
Cloud Configuration
Instance Configuration
Pricing Configuration
Choose your preferred pricing model
Enter your negotiated enterprise discount
AWS Enterprise Support fee percentage
Additional Storage
Personnel
Cloud Software Configuration
Calculating Costs
Updating summary...
Cloud Infrastructure
On-Premises Infrastructure
Comparison
Key Assumptions
Physical server counts follow the selected topology; cloud instance counts use only an exact GPU-to-SKU mapping or a manual quote.
Management software for compute and networking is included in the per unit cost.
IaaS pricing (AWS EC2) is gathered in real-time via AWS Pricing API for accurate, up-to-date costs.
SaaS Token API pricing (OpenAI, Anthropic, Gemini, Bedrock) is updated periodically based on published rates.
Personnel costs include benefits and overhead at 30% of base salary.
Power consumption of 10kW per GPU Server, 2.4kW for Control & Storage Servers. 2.4kW for Switches.
On-premises infrastructure has 24/7/365 100% utilization.
Token API pricing uses the input/output token ratio from GenAI Sizing inputs (avgInputTokens/avgOutputTokens) and requires GenAI Sizing calculator results.
Effective cost per 1M output tokens accounts for both compute-bound prefill and memory-bound decode phases using: (requests/sec × output_tokens/request), which accurately reflects workloads with varying input/output ratios.
Management network not included in the cost calculation.
B200 GPU AWS Instances assumed at the same price as H200. Pricing currently unavailable via API.
Roadmap
Launch v1.0!
2-4 GPU configurations per node support
Add cost per 1M token pricing
Token API SaaS pricing (OpenAI, Anthropic, Gemini, Bedrock)
Advanced networking cost modelling