Model catalog

Qwen2.5 7B Instruct

Qwen's 7B instruction-tuned model for users with more memory and a need for a broader general-purpose model.

Model overview

Available quantizations

Lower-bit quantizations generally use less memory. The compatibility calculator uses these candidates and its transparent approximate memory assumptions; it does not make a performance guarantee.

Compatibility notes

Weight estimates are approximate and include a runtime overhead. Real requirements vary with context length, runtime, drivers, and other system use. A model page cannot determine compatibility without your hardware profile; use the calculator for that assessment.

Qwen2.5 7B Instruct GGUF model card; verified value, last verified 2026-09-22. The GGUF publisher page documents a 32K context configuration; the base model advertises a longer context that is not assumed here. View source