Model catalog

DeepSeek-R1-Distill-Qwen 1.5B

DeepSeek's compact reasoning-focused model distilled from Qwen2.5, adding a smaller option for local reasoning experiments.

Model overview

Available quantizations

Lower-bit quantizations generally use less memory. The compatibility calculator uses these candidates and its transparent approximate memory assumptions; it does not make a performance guarantee.

Compatibility notes

Weight estimates are approximate and include a runtime overhead. Real requirements vary with context length, runtime, drivers, and other system use. A model page cannot determine compatibility without your hardware profile; use the calculator for that assessment.

DeepSeek-R1-Distill-Qwen 1.5B model card; verified value, last verified 2026-09-25. The publisher lists this distilled model under MIT and its config allows up to 131,072 tokens; its model card also notes the Qwen2.5 base-model license. GGUF candidates are tracked separately as a community conversion; runtime and long-context memory requirements are not verified. View source