TCO Analysis Dashboard

Directionally accurate total cost of ownership comparison between on-premises infrastructure, cloud, and SaaS AI APIs.

Calculating costs…

GPU Configuration

Use Case GPU Profiles

This requires 1 GPU servers

Affects both hardware pricing and cloud instance selection

GPU Configuration Summary

Use Case 1
8 GPUs
H2001 servers
p5en.48xlarge1 instances
Total (All Use Cases)
8 GPUs
1 GPU servers
8 GPUs per server

On-Premises Configuration

Hardware Breakdown

Calculated: 8 total GPUs ÷ 8 GPUs per server = 1 servers

GPU server power

System power per server. Enter your server’s rating, or use the platform default.

Auto-selected Ethernet network

$45,000

Sizes switches and connections for your deployment.

Single-homed access

Network details

One server uses its internal GPU interconnect; no external compute fabric is required. In-band and management connectivity are priced separately. Single-homed access; 48 endpoint ports per switch. Uplink and WAN/router ports are quote-dependent. Equipment costs assume new purchases. Shared-fabric allocation depends on the site.

Conventional HGX 400GbE compute fabric: 1 plane, 0 leaf and 0 spine switches. Includes converged in-band and out-of-band management; dedicated storage fabric is opt-in.

Converged in-band switches × 1$30,000
Converged in-band pluggables × 4$3,000
Out-of-band management switches × 1$12,000
Edit unit prices

Hardware Maintenance

Personnel

Facilities Configuration

Platform operating model

A few choices set the inferred platform, security, and observability footprint.

Platform operating model

Secure AI: Enterprise controls for governed AI applications and platforms.

Platform servicesAI security & governanceAudit-ready observability

Recommended topology: 3 control nodes. Production Kubernetes uses a three-node control plane; large clusters add a dedicated management node.

Included platform services

Add prices for the highlighted services to include them in totals.

Shared controls

Cisco AI DefenseAdd price· 1 application
ServiceOn-premisesCloud IaaS
Platform & OS
Commercial OS supportAdd price· 160 coresAdd price· 160 cores
Kubernetes platform supportAdd price· 3 nodesAdd price· 3 nodes
Security & AI governance
Isovalent EnterpriseAdd price· 3 nodesAdd price· 3 nodes
Observability
Splunk Observability Cloud$540/yr· 3 hosts$540/yr· 3 hosts

Cloud Configuration

Instance Configuration

Pricing Configuration

Choose your preferred pricing model

Enter your negotiated enterprise discount

AWS Enterprise Support fee percentage

Additional Storage

Personnel

Cloud Software Configuration

Calculating Costs

Updating summary...

Cloud Infrastructure

Total CostAdd price
Monthly Average
NPV

On-Premises Infrastructure

Total Cost$0.00
Monthly Average$0.00
NPV$0.00

Comparison

Key Assumptions

Physical server counts follow the selected topology; cloud instance counts use only an exact GPU-to-SKU mapping or a manual quote.

Management software for compute and networking is included in the per unit cost.

EC2 On-Demand prices come from AWS’s live pricing sources and are cached for up to 24 hours.

SaaS Token API pricing (OpenAI, Anthropic, Gemini, Bedrock) is updated periodically based on published rates.

Personnel costs include benefits and overhead at 30% of base salary.

GPU server power follows installed GPU TDP plus platform host overhead. Rack capacity uses peak power; training energy uses scheduled activity with a 30% idle-GPU allowance. Shared servers and switches use a 1.2 kW planning allowance each. PUE applies to electricity.

On-premises infrastructure has 24/7/365 100% utilization.

Token API pricing uses the input/output token ratio from GenAI Sizing inputs (avgInputTokens/avgOutputTokens) and requires GenAI Sizing calculator results.

Effective cost per 1M output tokens accounts for both compute-bound prefill and memory-bound decode phases using: (requests/sec × output_tokens/request), which accurately reflects workloads with varying input/output ratios.

Networking includes converged in-band and out-of-band management switches. Compute fabric follows the selected topology policy; dedicated storage fabric is opt-in. WAN uplinks and cables require site-specific quotes.

Each GPU uses its exact AWS instance price; unavailable or reserved prices require a matching quote.

Roadmap

Launch v1.0!

2-4 GPU configurations per node support

Add cost per 1M token pricing

Token API SaaS pricing (OpenAI, Anthropic, Gemini, Bedrock)

Advanced networking cost modelling