Back to models

Qwen 3.5 4B

Trending

Qwen/Qwen3.5-4B

A compact vision-language model that pairs strong reasoning with image understanding — light enough for a single 24GB GPU.

Qwen 3.5 4B is a multimodal vision-language model that accepts both text and images. At only 4B parameters it is the most affordable model in the catalog, booting on a single g5.xlarge while still offering up to 128K context on the large tier. Great for multimodal assistants, document understanding, and OCR-style workloads.

ProviderAlibaba Qwen
CategoryVision-Language
Parameters4B
PrecisionBF16
Context window128K
LicenseApache 2.0

Starting at

$1.01 / hour

g5.xlarge · NVIDIA A10G

You'll need a GPU Router account to deploy this model.

Min GPUA10G (24GB)
Params4B
Context128K