Skip to content

Meta · 3B · transformer-decoder

2024-09-25131K context3B params

Use Cases

chatcodemultilingualsummary

Quantization Options

QuantBitsVRAMQualityStatus
Q4_K_M43.3 GBModerate
Q8_0rec85.0 GBGood
F16168.0 GBExcellent

About this model

Llama 3.2 3B is a lightweight model from Meta designed for edge deployment and on-device inference. Despite its small size, it delivers surprisingly capable performance for text generation, summarization, and basic coding tasks. This model is ideal for users with limited hardware who still want a capable assistant. It runs comfortably on most modern laptops and even some mobile devices, making it one of the most accessible models in the Llama family.

Benchmarks

63.4
mmlu