Model catalog

Gemma 4 E4B IT

Google's instruction-tuned multimodal model has 8B total parameters and 4.5B effective parameters. Weight planning uses total parameters; the separate vision projector is excluded.

Model overview

Available quantizations

Lower-bit quantizations generally use less memory. The compatibility calculator uses these candidates and its transparent approximate memory assumptions; it does not make a performance guarantee.

Q4_0

The conversion repository lists gemma-4-E4B-it-Q4_0.gguf at 4.59 GB; its rounded decimal file size is converted to GiB for planning. This does not include context or runtime memory.

Bits per weight
4.5
Model size
4.28 GiB

Compatibility notes

Estimates use the cataloged model-weight size (or a parameter-based estimate), apply weight overhead, and include standard runtime overhead. They do not include context/KV-cache memory. Real requirements vary with runtime, drivers, and other system use. A model page cannot determine compatibility without your hardware profile; use the calculator for that assessment.

Google Gemma 4 E4B IT model card; verified value, last verified 2026-09-29. Google identifies 8B total and 4.5B effective parameters, Apache-2.0 licensing, and up to 128K context. The GGUF conversion has a separate vision-projector file, excluded from the listed main-weight size. View source

Multimodal memory scope: This estimate excludes the separate vision/projector file and its runtime memory. ggml-org Gemma 4 E4B IT GGUF repository