LLM Hardware & Cost Architect
Open-source LLM calculator

Plan inference & fine-tuning — hardware, VRAM and cost.

Compare open-source LLMs across a catalog of hundreds of models; live cloud prices, TCO and per-token cost analysis. No sign-up, free.

Start Calculating →

Real scenarios, real numbers

These cards are computed server-side by the same calculator engine; each is editable in the app.

Inference

TTFT, TPOT, tokens/s and VRAM — from 8B to 671B MoE.

Fine-Tuning

GPU hours, VRAM and platform cost for QLoRA, LoRA and full fine-tuning.

Live Cloud Prices

Up-to-date GPU prices from RunPod, Lambda and Modal; on-prem TCO comparison.

125 open-source models in the catalog