Skip to content

Meta · 1B · transformer-decoder

2024-09-25131K context1B params

Use Cases

chatsummary

Quantization Options

QuantBitsVRAMQualityStatus
Q4_K_M42.1 GBModerate
Q8_0rec83.0 GBGood
F16164.0 GBExcellent

About this model

Llama 3.2 1B is the smallest model in the Llama 3.2 family, designed for ultra-lightweight deployment scenarios. It can handle basic text generation and summarization tasks while requiring minimal compute resources. This model is best suited for simple tasks, prototyping, or situations where hardware is extremely constrained. It runs on virtually any modern device and provides fast inference even on CPU-only setups.

Benchmarks

49.3
mmlu