Q4_K_M
Common lower-memory GGUF choice; actual files are produced by community or vendor conversion pipelines.
- Bits per weight
- 4.5
Model catalog
Qwen's 7B instruction-tuned model for users with more memory and a need for a broader general-purpose model.
Lower-bit quantizations generally use less memory. The compatibility calculator uses these candidates and its transparent approximate memory assumptions; it does not make a performance guarantee.
Common lower-memory GGUF choice; actual files are produced by community or vendor conversion pipelines.
Middle-ground GGUF choice with a larger memory requirement than Q4_K_M.
Higher-memory GGUF choice; it is still an approximate memory-planning input rather than a performance claim.
Weight estimates are approximate and include a runtime overhead. Real requirements vary with context length, runtime, drivers, and other system use. A model page cannot determine compatibility without your hardware profile; use the calculator for that assessment.
Qwen2.5 7B Instruct GGUF model card; verified value, last verified 2026-09-22. The GGUF publisher page documents a 32K context configuration; the base model advertises a longer context that is not assumed here. View source