GeForce RTX 3060 12GB
A 12 GiB discrete GPU that leaves room for many small and medium quantized models, subject to system RAM and context requirements.
- Vendor
- NVIDIA
- Type
- Discrete
- Dedicated VRAM
- 12 GiB
- Architecture
- Ampere
Curated GPU catalog
Compare the catalog’s dedicated memory capacities and documented hardware details. Open a GPU for its source and planning notes.
A 12 GiB discrete GPU that leaves room for many small and medium quantized models, subject to system RAM and context requirements.
An 8 GiB discrete GPU for smaller and medium quantized models; larger models may require partial offload or CPU execution.
A 12 GiB discrete GPU with the same capacity class as other 12 GiB cards; the calculator intentionally does not infer speed from the GPU name.
A 24 GiB discrete GPU with substantial capacity for larger quantized models, without implying a particular speed or context guarantee.
An 8 GiB discrete GPU; actual llama.cpp backend support and practical results depend on the selected runtime and driver stack.
A 16 GiB discrete GPU with capacity for a wider range of quantized models; compatibility still depends on runtime support and system memory.
A 24 GiB discrete GPU with high memory capacity; no backend performance or universal compatibility claim is made.
An 8 GiB discrete GPU; backend and driver support should be checked for the chosen llama.cpp build.
A 16 GiB discrete GPU; local LLM results depend on the llama.cpp backend, driver, and offload configuration.
Integrated graphics with no dedicated VRAM entry; LLMGauge conservatively treats shared system memory as CPU-only capacity.
An older 6 GiB discrete card; useful for evaluating smaller quantized models, with limited dedicated-memory headroom.
The 16 GiB RTX 4060 Ti variant offers more model-memory capacity than the 8 GiB version; the catalog treats capacity separately from speed.
A 16 GiB discrete GPU that broadens the catalog's mid/high capacity range without implying model speed or guaranteed fit.
A current 16 GiB discrete option; compatibility uses memory capacity only and does not infer backend availability or performance.
An established 8 GiB discrete GPU; actual llama.cpp backend and driver support depend on the user's software stack.
A 12 GiB RDNA 3 card that fills the catalog's AMD mid-capacity range; memory fit does not confirm backend support.
A newer 16 GiB discrete option; LLMGauge reports its memory class while leaving runtime and driver compatibility to the user's setup.
The RX 9070 XT provides a current 16 GiB AMD option; no speed or universal llama.cpp support is inferred.
A 10 GiB Intel discrete GPU; check the selected llama.cpp build and driver for backend support before relying on offload.
A 12 GiB Intel discrete GPU that expands the memory-capacity range; available runtime paths remain software-dependent.
Integrated graphics use system memory rather than a fixed dedicated VRAM capacity; LLMGauge conservatively classifies this as CPU-only capacity.