Text / Qwen

Qwen3.5 · 9B

An open-weight vision-language model with a compact Ollama package suitable for a private assistant or coding help.

Verified sourceReleased Feb 27, 2026Source checked 9/23/2026
LOCALRENTED GPUOPEN WEIGHTS

At a glance

Parameters
9B
Architecture
Hybrid attention
Context length
262,144 native
License
Apache 2.0
Disk space
6.6 GB
Software
Ollama, MLX, Transformers, vLLM
View the model source
MY HARDWARE

Will it run on your machine?

Save your machine to see a personalized rating and its reasoning.

Add my hardware

Best for

  • Private chat and coding help on a single computer.
  • Image understanding when the chosen runtime and memory support it.

Tradeoffs

  • Vision input and very long context need separate memory validation.

Ways to run it

Ollama · Windows, macOS, Linux

Ollama lists a 6.6 GB package. Text-only memory estimates do not cover vision input.

ollama run qwen3.5:9b
Official instructions

MLX · macOS

ollama run qwen3.5:9b-mlx
Official instructions

Explore its uses

CHATCODINGREASONINGMULTIMODAL

Get it running

Set up a private AI chat assistant

Install a local runtime, run Qwen3.5 9B, confirm responses, and know when to choose the smaller 4B package.

View workflow

Ask your own documents with local AI

Pair Qwen3.5 in Ollama with Open WebUI’s document retrieval, then check that answers cite your uploaded file.

View workflow

Keep exploring