Back to models
Qwen 3.5 4B
TrendingQwen/Qwen3.5-4B
A compact vision-language model that pairs strong reasoning with image understanding — light enough for a single 24GB GPU.
Qwen 3.5 4B is a multimodal vision-language model that accepts both text and images. At only 4B parameters it is the most affordable model in the catalog, booting on a single g5.xlarge while still offering up to 128K context on the large tier. Great for multimodal assistants, document understanding, and OCR-style workloads.
ProviderAlibaba Qwen
CategoryVision-Language
Parameters4B
PrecisionBF16
Context window128K
LicenseApache 2.0
Starting at
$1.01 / hour
g5.xlarge · NVIDIA A10G
You'll need a GPU Router account to deploy this model.
Min GPUA10G (24GB)
Params4B
Context128K