Q4_K_M
Common lower-memory GGUF choice; actual files are produced by community or vendor conversion pipelines.
- Bits per weight
- 4.5
Model catalog
Mistral's 7B instruction-tuned model with a broad ecosystem of local GGUF conversions.
Lower-bit quantizations generally use less memory. The compatibility calculator uses these candidates and its transparent approximate memory assumptions; it does not make a performance guarantee.
Common lower-memory GGUF choice; actual files are produced by community or vendor conversion pipelines.
Middle-ground GGUF choice with a larger memory requirement than Q4_K_M.
Higher-memory GGUF choice; it is still an approximate memory-planning input rather than a performance claim.
Weight estimates are approximate and include a runtime overhead. Real requirements vary with context length, runtime, drivers, and other system use. A model page cannot determine compatibility without your hardware profile; use the calculator for that assessment.
Mistral 7B Instruct v0.3 model card; verified value, last verified 2026-09-22. The public model card identifies the instruct variant; this catalog uses a conservative 32K maximum for local GGUF planning. View source