Skip to content

Qwen 3.5 4B

Apache 2.0

Alibaba · 4B · transformer-decoder

2026-03-02262K context4B params

Use Cases

chatcodereasoningmultilingualvisiontools

Quantization Options

QuantBitsVRAMQualityStatus
Q4_K_Mrec44.5 GBGood
Q8_086.5 GBGood

About this model

Qwen 3.5 4B is an efficient multimodal model with strong reasoning and tool-use capabilities. It supports thinking and non-thinking modes, letting you trade speed for reasoning depth. At 4.5 GB VRAM (Q4) it fits on 8 GB GPUs with headroom, making it a practical daily driver for local AI on modest hardware.