Q4_K_M
Common lower-memory GGUF choice; actual files are produced by community or vendor conversion pipelines.
- Bits per weight
- 4.5
Model catalog
Google's compact instruction-tuned Gemma 3 text model for constrained local systems.
Lower-bit quantizations generally use less memory. The compatibility calculator uses these candidates and its transparent approximate memory assumptions; it does not make a performance guarantee.
Common lower-memory GGUF choice; actual files are produced by community or vendor conversion pipelines.
Middle-ground GGUF choice with a larger memory requirement than Q4_K_M.
Higher-memory GGUF choice; it is still an approximate memory-planning input rather than a performance claim.
Weight estimates are approximate and include a runtime overhead. Real requirements vary with context length, runtime, drivers, and other system use. A model page cannot determine compatibility without your hardware profile; use the calculator for that assessment.
Google Gemma 3 model card; verified value, last verified 2026-09-22. The publisher describes the Gemma 3 family as supporting a 128K context window; runtime and conversion support may impose lower practical limits. View source