Model catalog

Qwen3 8B

An 8B Qwen3 general-purpose model that adds a mid-sized option beyond the existing Qwen2.5 entries.

Model overview

Available quantizations

Lower-bit quantizations generally use less memory. The compatibility calculator uses these candidates and its transparent approximate memory assumptions; it does not make a performance guarantee.

Compatibility notes

Weight estimates are approximate and include a runtime overhead. Real requirements vary with context length, runtime, drivers, and other system use. A model page cannot determine compatibility without your hardware profile; use the calculator for that assessment.

Qwen3 8B publisher model card; verified value, last verified 2026-09-23. Publisher model metadata and license. GGUF candidates come from a separate Qwen conversion repository; extended YaRN context is not represented as the default maximum here. View source