Skip to content

Meta · 8B · transformer-decoder

2024-07-23131K context8B params

Use Cases

chatcodemultilingualtoolssummary

Quantization Options

QuantBitsVRAMQualityStatus
Q4_K_M46.3 GBModerate
Q8_0rec810.0 GBGood
F161618.0 GBExcellent

About this model

Llama 3.1 8B is Meta's versatile mid-range model offering strong performance across text generation, coding, and multilingual tasks. It supports the full 128K context window and includes native tool-use capabilities. This model strikes an excellent balance between capability and resource requirements, running well on consumer GPUs with 8-16GB VRAM. It is one of the most popular open-source models for local inference and serves as a strong baseline for many use cases.

Benchmarks

73.0
mmlu