Q4_K_M
Common lower-memory GGUF choice; actual files are produced by community or vendor conversion pipelines.
- Bits per weight
- 4.5
Model catalog
Google's 4B instruction-tuned Gemma 3 model with text and image-input capabilities in its published model family.
Lower-bit quantizations generally use less memory. The compatibility calculator uses these candidates and its transparent approximate memory assumptions; it does not make a performance guarantee.
Common lower-memory GGUF choice; actual files are produced by community or vendor conversion pipelines.
Middle-ground GGUF choice with a larger memory requirement than Q4_K_M.
Higher-memory GGUF choice; it is still an approximate memory-planning input rather than a performance claim.
Weight estimates are approximate and include a runtime overhead. Real requirements vary with context length, runtime, drivers, and other system use. A model page cannot determine compatibility without your hardware profile; use the calculator for that assessment.
Google Gemma 3 model card; verified value, last verified 2026-09-22. The calculator models text-generation memory only; it does not estimate image-encoder memory or multimodal runtime behavior. View source