System ready
Send Feedback

TCO Analysis Dashboard

Directionally accurate total cost of ownership comparison between on-premises infrastructure, cloud, and SaaS AI APIs.

TCO Comparison: Cloud vs. On-Premises

3 YEARS ANALYSIS

Cloud Pricing Model

Analysis Period

Calculating TCO Analysis

Fetching cloud pricing data and computing cost comparisons for your configuration...

On-Premises Total Cost

$2,527,591.00

$70,210.86/month avg.

Cloud Total Cost

$246,227.00

$6,839.64/month avg.

Recommended Option

CLOUD

Cloud infrastructure is more cost-effective for your use case, saving approximately $2,281,364.00 (90.3%) over 3 years.

Savings
90.3%
Payback Period
Beyond analysis period
On-Premises Cost Breakdown
Compute
1x GPU Servers • 1x Control Servers
(1 × $400,000) + (1 × $30,000)
$0.00
Network
Derived Ethernet BOM • compute, in-band, and OOB
Fabric-specific switch and optic rates
$0.00
Storage
0x Storage Servers
0 × $30,000
$0.00
Power & Cooling
24/7/365 Operation • PUE: 1.5
1.0kW × 8,760 hrs/yr × 1.5 PUE × $0.120/kWh × 3 yrs
$0.00
IT Labor & Admin
2x Platform Engineers • 1x Network Engineers • 30% Benefits
[(2 × $150,000) + (1 × $130,000)] × 1.3 (w/ benefits) × 3 yrs
$0.00
Facilities
Data Center Rack Space (2 racks, limited by power)
2 racks × $1,200/rack/month × 12 months/yr × 3 yrs • 30kW per rack
$0.00
Maintenance & Support
Hardware & Software Support
$0.00 (hardware) × 10% × 3 yrs
$0.00
Software
Inferred platform services
Quote-backed platform services are added when configured
$0.00
Total$0.00
Cloud Cost Breakdown
EC2 Compute (p5en.48xlarge)
1x instances • on-demand
On-Demand rate: $0.000/hour per instance
Effective rate: $0.000/hour per instance
$0.00
EBS Storage (gp3)
0GB • General Purpose SSD
On-Demand rate: $0.080/GB/month
0
$0.00
Enterprise Support
AWS Enterprise Support Plan • 10% of net AWS usage
10% × $0.00 (net AWS usage)
$0.00
Cloud Architecture & DevOps
2x Cloud Engineers • 30% Benefits
2 engineers × $180,000/year × 1.3 (w/ benefits) × 3 years
$0.00
Total$0.00

GPU Configuration

Use Case GPU Profiles

This requires 1 GPU servers

Affects both hardware pricing and cloud instance selection

GPU Configuration Summary

Use Case 1
8 GPUs
H2001 servers
p5en.48xlarge1 instances
Total (All Use Cases)
8 GPUs
1 GPU servers
8 GPUs per server

On-Premises Configuration

Hardware Breakdown

Calculated: 8 total GPUs ÷ 8 GPUs per server = 1 servers

Auto-selected Ethernet network

$1,226,000

Conventional HGX 400GbE compute fabric: 1 plane, 2 leaf and 2 spine switches. Includes converged in-band and out-of-band management; dedicated storage fabric is opt-in.

400GbE compute backend switches (2 leaf + 2 spine) × 4$440,000
400GbE compute endpoint pluggables × 16$56,000
400GbE compute fabric pluggables × 256$640,000
Converged in-band switches × 2$60,000
Converged in-band pluggables × 8$6,000
Out-of-band management switches × 2$24,000
Adjust quote rates

Hardware Maintenance

Personnel

Facilities Configuration

Platform operating model

A few choices set the inferred platform, security, and observability footprint.

Platform operating model

Secure AI: Enterprise controls for governed AI applications and platforms.

Platform servicesAI security & governanceAudit-ready observability

Recommended topology: 3 control nodes. Production Kubernetes uses a three-node control plane; large clusters add a dedicated management node.

Included platform services

Quantities are inferred. Quote-required items are visible but excluded from totals until priced.

Shared controls

Cisco AI DefenseQuote required· 1 application
ServiceOn-premisesCloud IaaS
Platform & OS
Commercial OS supportQuote required· 160 coresQuote required· 160 cores
Kubernetes platform supportQuote required· 3 nodesQuote required· 3 nodes
Security & AI governance
Isovalent EnterpriseQuote required· 3 nodesQuote required· 3 nodes
Observability
Splunk Observability Cloud$540/yr· 3 hosts$540/yr· 3 hosts

Cloud Configuration

Instance Configuration

Pricing Configuration

Choose your preferred pricing model

Enter your negotiated enterprise discount

AWS Enterprise Support fee percentage

Additional Storage

Personnel

Cloud Software Configuration

Calculating Costs

Updating summary...

Cloud Infrastructure

Total Cost$0.00
Monthly Average$0.00
NPV$0.00

On-Premises Infrastructure

Total Cost$0.00
Monthly Average$0.00
NPV$0.00

Comparison

Cost Difference$0.00
Savings Percentage0.0%
Break-even PointBeyond analysis period

Key Assumptions

Physical server counts follow the selected topology; cloud instance counts use only an exact GPU-to-SKU mapping or a manual quote.

Management software for compute and networking is included in the per unit cost.

IaaS pricing (AWS EC2) is gathered in real-time via AWS Pricing API for accurate, up-to-date costs.

SaaS Token API pricing (OpenAI, Anthropic, Gemini, Bedrock) is updated periodically based on published rates.

Personnel costs include benefits and overhead at 30% of base salary.

Power consumption of 10kW per GPU Server, 2.4kW for Control & Storage Servers. 2.4kW for Switches.

On-premises infrastructure has 24/7/365 100% utilization.

Token API pricing uses the input/output token ratio from GenAI Sizing inputs (avgInputTokens/avgOutputTokens) and requires GenAI Sizing calculator results.

Effective cost per 1M output tokens accounts for both compute-bound prefill and memory-bound decode phases using: (requests/sec × output_tokens/request), which accurately reflects workloads with varying input/output ratios.

Management network not included in the cost calculation.

B200 GPU AWS Instances assumed at the same price as H200. Pricing currently unavailable via API.

Roadmap

Launch v1.0!

2-4 GPU configurations per node support

Add cost per 1M token pricing

Token API SaaS pricing (OpenAI, Anthropic, Gemini, Bedrock)

Advanced networking cost modelling