AI & Local
LLM Memory Calculator
Two estimates in one place. Running a model: weights at your chosen quantisation, the KV cache for your context length, and runtime overhead. Fine-tuning it: frozen base, gradients, optimizer state and activations for full, LoRA or QLoRA training, so you can see why a model that runs fine will not fine-tune on the same machine.
Free · runs in your browser · no sign-up to use.
Assumes 0.5 bytes/param for weights and a 2-byte KV cache, plus ~18% runtime overhead. KV is estimated from the parameter count.
Estimated memory to run
Weights3.73 GB
KV cache4.25 GB
Runtime overhead1.44 GB
Total9.42 GB
Estimate, not a guarantee. Actual use varies with the runtime, batching, and quant format.
Related tools
Free, and they run in your browser too.
- AI & LocalWhat LLM Can My Mac Run?See which local LLMs fit your Mac's memory.Open
- AI & LocalAI Context PackBuild reusable personal context files for AI assistants.Open
- AI & LocalSpeaker KitTurn your profile into a polished speaker one-sheet and bios.Open
- AI & LocalPPE Photo CheckerCheck a site photo for hard hats, vests and masks.Open