Qwen3.5 9B
Qwen3.5 9B is the small unified vision-language model in the Qwen3.5 series (Apache 2.0, February 2026): a dense 9B network that reads text, images and video. It covers reasoning, coding, agents and visual understanding at low inference cost, thinks by default and lets you switch that off. The series claims coverage of 201 languages and dialects — a practical first model for extracting data from scans and screenshots. It is fulfilled through global providers under our European contract, and requests go only to backends that do not retain prompts or completions.
model id: qwen3.5-9b
from openai import OpenAI client = OpenAI( api_key="hb-...", base_url="https://api.heabsy.com/v1",) r = client.chat.completions.create( model="qwen3.5-9b", messages=[{"role": "user", "content": "Hello"}],)print(r.choices[0].message.content)
curl https://api.heabsy.com/v1/chat/completions \ -H "Authorization: Bearer $HEABSY_API_KEY" \ -H "Content-Type: application/json" \ -d '{ "model": "qwen3.5-9b", "messages": [{"role": "user", "content": "Hello"}] }'
import OpenAI from "openai"; const client = new OpenAI({ apiKey: process.env.HEABSY_API_KEY, baseURL: "https://api.heabsy.com/v1",}); const r = await client.chat.completions.create({ model: "qwen3.5-9b", messages: [{ role: "user", content: "Hello" }],});console.log(r.choices[0].message.content);
Price list, updated 12 September 2026. No minimum spend, no subscription.
About the model
- Developer
- Qwen (Alibaba)
- Released
- February 2026
- Licence
- Apache-2.0
- Architecture
- dense · 9B parameters
- Input
- text, images, video
- Reasoning
- switchable per request
- Languages named by the developer
- 201 languages
Source: the developer's model card, as of 17 September 2026.
Licence: what it means for your company
Apache-2.0. A permissive open-source licence: you may use, modify and sell products built on the model, including commercially; when you redistribute the weights, the licence and copyright notices go with them.
A summary of the licence, not legal advice.
Technical details
- Context window
- 256K tokens
- Max output
- 65,536 tokens
- Throughput limit
- 200K tok/min
- Tokenizer
- Other
- Streaming
- yes
- Open weights
- qwen/qwen3.5-9b
- max_tokens
- 1 … 65536
- temperature
- 0 … 2
- top_p
- 0 … 1
- frequency_penalty
- -2 … 2
- presence_penalty
- -2 … 2
- stop
- array
- seed
- integer
This model is fulfilled through global providers under our European contract: one DPA with an EU company, one invoice for every model in the catalog, no US counterparty on your paperwork. Compute may run outside the EEA — if you need strict EEA-only processing, use our own-hardware model.
Questions about this model
›Where is Qwen3.5 9B computed?
Outside the EEA, through a global provider under our European contract. You get one DPA with an EU company and one invoice, but the compute itself does not run in the EEA — we state that instead of hiding it.
›Does Qwen3.5 9B store prompts and completions?
No. This model runs with zero data retention: request content is not written to durable storage and nothing is used for training.
›How large is the Qwen3.5 9B context window?
256K tokens, with up to 65,536 tokens of output.
›What does Qwen3.5 9B cost?
$0.15 per million input tokens and $0.23 per million output tokens. No minimum spend and no subscription; the same price list covers the API and the invoice.
›Is there a throughput limit on Qwen3.5 9B?
Yes, 200K tokens per minute. Higher limits are a question of volume, not of principle.
›Which languages does Qwen3.5 9B support?
The developer states support for 201 languages without listing them. Whether your language is among them is not documented — test on your own material first.
›Under which licence is Qwen3.5 9B released?
Apache-2.0. A permissive open-source licence: you may use, modify and sell products built on the model, including commercially; when you redistribute the weights, the licence and copyright notices go with them.
›Can reasoning be switched off in Qwen3.5 9B?
Yes. Per the developer, thinking can be switched on or off for each request.
Compare with similar models
| Model | Developer | Parameters | Context | Licence | Input / 1M |
|---|---|---|---|---|---|
| Qwen3.5 9B | Qwen (Alibaba) | 9B | 256K | Apache-2.0 | $0.15 |
| Qwen3 32B | Qwen (Alibaba) | 32.8B | 128K | Apache-2.0 | $0.12 |
| Qwen3.6 35B A3B | Qwen (Alibaba) | 35B / 3B | 256K | Apache-2.0 | $0.15 |
| Qwen3 Next 80B A3B Instruct | Qwen (Alibaba) | 80B / 3B | 256K | Apache-2.0 | $0.14 |
30 more models behind the same key and the same invoice.
Browse the catalog