Q4_0
The conversion repository lists gemma-4-12B-it-Q4_0.gguf at 7.22 GB; its rounded decimal file size is converted to GiB for planning. This does not include context or runtime memory.
- Bits per weight
- 4.5
- Model size
- 6.72 GiB
Model catalog
Google's instruction-tuned 11.95B multimodal model. Listed GGUF sizing covers the main model weights; the separate vision projector is excluded.
Lower-bit quantizations generally use less memory. The compatibility calculator uses these candidates and its transparent approximate memory assumptions; it does not make a performance guarantee.
The conversion repository lists gemma-4-12B-it-Q4_0.gguf at 7.22 GB; its rounded decimal file size is converted to GiB for planning. This does not include context or runtime memory.
Estimates use the cataloged model-weight size (or a parameter-based estimate), apply weight overhead, and include standard runtime overhead. They do not include context/KV-cache memory. Real requirements vary with runtime, drivers, and other system use. A model page cannot determine compatibility without your hardware profile; use the calculator for that assessment.
Google Gemma 4 12B IT model card; verified value, last verified 2026-09-29. Google documents an 11.95B multimodal instruction model, Apache-2.0 licensing, and up to 256K context. The selected GGUF conversion lists the main model file separately from its auxiliary vision projector. View source
Multimodal memory scope: This estimate excludes the separate vision/projector file and its runtime memory. ggml-org Gemma 4 12B IT GGUF repository