Model catalog

Llama 3.1 8B Instruct

Meta's 8B instruction-tuned Llama 3.1 model with publisher-documented long-context metadata.

Model overview

Available quantizations

Lower-bit quantizations generally use less memory. The compatibility calculator uses these candidates and its transparent approximate memory assumptions; it does not make a performance guarantee.

Compatibility notes

Weight estimates are approximate and include a runtime overhead. Real requirements vary with context length, runtime, drivers, and other system use. A model page cannot determine compatibility without your hardware profile; use the calculator for that assessment.

Meta Llama 3.1 8B Instruct model card; verified value, last verified 2026-09-23. Publisher card reports 8B parameters and 128K context. GGUF conversion is community-provided and remains subject to the model's community license. View source