Model catalog

Qwen3-VL-8B-Instruct

Qwen's 8B-class image and video model is listed with 9B total parameters for weight planning. Its vision projector and multimodal runtime memory are not estimated separately.

Model overview

Available quantizations

Lower-bit quantizations generally use less memory. The compatibility calculator uses these candidates and its transparent approximate memory assumptions; it does not make a performance guarantee.

Q4_K_M

The conversion repository lists Qwen3VL-8B-Instruct-Q4_K_M.gguf at 5.03 GB; its rounded decimal file size is converted to GiB for planning. This does not include context or runtime memory.

Bits per weight
4.5
Model size
4.68 GiB

Compatibility notes

Estimates use the cataloged model-weight size (or a parameter-based estimate), apply weight overhead, and include standard runtime overhead. They do not include context/KV-cache memory. Real requirements vary with runtime, drivers, and other system use. A model page cannot determine compatibility without your hardware profile; use the calculator for that assessment.

Qwen3-VL-8B-Instruct publisher model card; verified value, last verified 2026-10-01. Qwen identifies the instruct variant as an 8B-class multimodal model; its Hugging Face model card lists 9B parameters, used here for total-weight planning. Qwen documents Apache-2.0 and 256K context. View source

Multimodal memory scope: This estimate excludes the separate vision/projector file and its runtime memory. Qwen Qwen3-VL-8B-Instruct GGUF repository