Model catalog

Gemma 3n E4B IT

Google's E4B effective-size multimodal model contains 8B total parameters. Weight planning uses the total count; multimodal runtime memory is not separately estimated.

Model overview

Available quantizations

Lower-bit quantizations generally use less memory. The compatibility calculator uses these candidates and its transparent approximate memory assumptions; it does not make a performance guarantee.

Q8_0

The conversion repository lists gemma-3n-E4B-it-Q8_0.gguf at 7.35 GB; its rounded decimal file size is converted to GiB for planning. This does not include context or runtime memory.

Bits per weight
8
Model size
6.85 GiB

Compatibility notes

Estimates use the cataloged model-weight size (or a parameter-based estimate), apply weight overhead, and include standard runtime overhead. They do not include context/KV-cache memory. Real requirements vary with runtime, drivers, and other system use. A model page cannot determine compatibility without your hardware profile; use the calculator for that assessment.

Google Gemma 3n model card; verified value, last verified 2026-09-30. Google documents E4B as an effective-size label distinct from the model's 8B total stored parameters, multimodal text/image/video/audio inputs, and 32K context. Google lists Gemma Terms of Use; multimodal runtime memory is not separately estimated. View source

Multimodal memory scope: The catalog has not verified whether the selected GGUF weights include all multimodal components or require separate files. Multimodal runtime memory is not estimated and may be missing from this result. ggml-org Gemma 3n E4B IT GGUF repository