Gemma 3 12B IT
The 12B Gemma 3 instruction model adds a larger text-and-image family option; this estimate covers text weights only.
- Provider
- Family
- Gemma 3
- Parameters
- 12B
- Formats
- gguf, safetensors
Curated model catalog
Compare cataloged model families, parameter sizes, and recorded formats. Open a model for its quantization candidates, sources, and compatibility notes.
20 models found.
The 12B Gemma 3 instruction model adds a larger text-and-image family option; this estimate covers text weights only.
Google's compact instruction-tuned Gemma 3 text model for constrained local systems.
Google's largest Gemma 3 instruction variant, useful for exploring high-memory local configurations.
Google's 4B instruction-tuned Gemma 3 model with text and image-input capabilities in its published model family.
Meta's 8B instruction-tuned Llama 3.1 model with publisher-documented long-context metadata.
Meta's compact instruction-tuned Llama model for lightweight local text generation.
Meta's 3B instruction-tuned Llama model, suitable for modest local hardware when quantized.
Meta's 70B text-only instruction model, extending the catalog into the high-memory class.
Mistral's 7B instruction-tuned model with a broad ecosystem of local GGUF conversions.
A 12B Mistral and NVIDIA collaboration model, extending the catalog into a larger multilingual instruction tier.
Mistral's 24B instruction-tuned model; the memory estimate covers text weights and not its image-processing path.
Microsoft's compact Phi instruction model with a documented long-context capability.
Microsoft's 3.8B instruction model provides a compact alternative with a publisher-documented 128K context limit.
Qwen's 3B instruction-tuned model with a published 32K context configuration for its GGUF variant.
Qwen's 7B instruction-tuned model for users with more memory and a need for a broader general-purpose model.
A code-focused instruction model, adding a distinct software-development use case to the catalog.
Qwen's 14.8B general-purpose Qwen3 model, filling the size range between the existing 8B and 27B entries.
Qwen's 32.8B general-purpose Qwen3 model for users evaluating the larger end of the local model range.
Qwen's compact 4B general-purpose model; a separate GGUF conversion is listed for llama.cpp workflows.
An 8B Qwen3 general-purpose model that adds a mid-sized option beyond the existing Qwen2.5 entries.