Skip to content

GPU Calc

LLM inference sizing, GPU comparison, and cost modeling for engineers and infrastructure teams.

Quick Estimate

Fast GPU memory and throughput estimate from model size and serving parameters.


Open tool →
Advanced Calculator

Detailed inference sizing with batching, quantization, KV cache, and cost modeling.


Open tool →
GPU Explorer

Compare GPUs across memory, throughput, cost, and availability tiers.


Open tool →
Hybrid Savings

Model cost savings between cloud, on-premise, and hybrid GPU deployment strategies.


Open tool →
Routing Economics

Analyze request routing between model tiers to optimize cost vs quality tradeoffs.


Open tool →