Work in progress

GPU Calc

LLM inference sizing, GPU comparison, and cost modeling for engineers and infrastructure teams.

KV Cache Calculator

Calculate KV cache memory requirements for any model on supported GPU systems.


Open tool →
Advanced Calculator

Detailed inference sizing with batching, quantization, KV cache, and cost modeling.


Open tool →
GPU Explorer

Compare GPUs across memory, throughput, cost, and availability tiers.


Open tool →
Hybrid Savings

Model cost savings between cloud, on-premise, and hybrid GPU deployment strategies.


Open tool →
Routing Economics

Analyze request routing between model tiers to optimize cost vs quality tradeoffs.


Open tool →