Hardware basics
What Is VRAM?
A practical explanation of GPU memory and why it matters when running local language models.
Read this guideLLMGauge guides
Short, source-aware explanations of the memory and runtime concepts behind the calculator.
Hardware basics
A practical explanation of GPU memory and why it matters when running local language models.
Read this guideModel basics
Understand how lower-precision weights reduce memory use and introduce quality trade-offs.
Read this guideRuntime basics
See what it means to run a local model entirely on a GPU, partly on a GPU, or on the CPU.
Read this guideRuntime basics
Why the amount of text a model can consider affects memory beyond the weight file.
Read this guideUsing LLMGauge
Understand what the calculator checks, what its result categories mean, and where uncertainty remains.
Read this guide