Q4_K_M
The conversion repository lists Qwen3VL-8B-Instruct-Q4_K_M.gguf at 5.03 GB; its rounded decimal file size is converted to GiB for planning. This does not include context or runtime memory.
- Bits per weight
- 4.5
- Model size
- 4.68 GiB
Model catalog
Qwen's 8B-class image and video model is listed with 9B total parameters for weight planning. Its vision projector and multimodal runtime memory are not estimated separately.
Lower-bit quantizations generally use less memory. The compatibility calculator uses these candidates and its transparent approximate memory assumptions; it does not make a performance guarantee.
The conversion repository lists Qwen3VL-8B-Instruct-Q4_K_M.gguf at 5.03 GB; its rounded decimal file size is converted to GiB for planning. This does not include context or runtime memory.
Estimates use the cataloged model-weight size (or a parameter-based estimate), apply weight overhead, and include standard runtime overhead. They do not include context/KV-cache memory. Real requirements vary with runtime, drivers, and other system use. A model page cannot determine compatibility without your hardware profile; use the calculator for that assessment.
Qwen3-VL-8B-Instruct publisher model card; verified value, last verified 2026-10-01. Qwen identifies the instruct variant as an 8B-class multimodal model; its Hugging Face model card lists 9B parameters, used here for total-weight planning. Qwen documents Apache-2.0 and 256K context. View source
Multimodal memory scope: This estimate excludes the separate vision/projector file and its runtime memory. Qwen Qwen3-VL-8B-Instruct GGUF repository