Qwen: Qwen3.5-9B

qwen/qwen3.5-9b

Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design with early fusion of multimodal tokens, allowing the model to process and reason across text and images within the same context.

Model specifications

Input
text, image, video
Output
text
Context
262,144 tokens
Max output
262,144 tokens
Input price
$0.04 / 1M tokens
Output price
$0.15 / 1M tokens
Released
2026-03-11

Capabilities

  • Streaming
  • Function calling
  • Vision
  • JSON mode
  • Playground

Provider pricing, discounts and data privacy

Compare effective provider prices, published discounts, regions, retention policies, training use, compliance, and privacy links by service tier.

Standard service tier

2 available providers · tier input average $0.105 / 1M tokens · tier output average $0.2 / 1M tokens

DeepInfra

Tier: Standard · Region: US · Quantization: fp16

Pricing
Input
$0.04 / 1M tokens
Output
$0.15 / 1M tokens

No provider discount is currently published.

Data privacy and compliance
Region
US
Zero data retention
Yes
Data retention
Zero retention
Used for training
No
Data collection
Moderated
No
GDPR compliant
Yes
HIPAA compliant
No
SOC 2 certified
Yes
BYOK supported
Yes

Together

Tier: Standard · Region: US · Quantization: fp8

Pricing
Input
$0.17 / 1M tokens
Output
$0.25 / 1M tokens

No provider discount is currently published.

Data privacy and compliance
Region
US
Zero data retention
No
Data retention
Unknown retention
Used for training
No
Data collection
Moderated
No
GDPR compliant
No
HIPAA compliant
No
SOC 2 certified
No
BYOK supported
No

Frequently asked questions

What is Qwen: Qwen3.5-9B?
Qwen3.5-9B is a multimodal foundation model from the Qwen3.5 family, designed to deliver strong reasoning, coding, and visual understanding in an efficient 9B-parameter architecture. It uses a unified vision-language design with early fusion of multimodal tokens, allowing the model to process and reason across text and images within the same context.
How much does Qwen: Qwen3.5-9B cost?
Input costs start at $0.04 / 1M tokens and output costs start at $0.15 / 1M tokens. Provider-level prices vary by service tier.
What is the context length of Qwen: Qwen3.5-9B?
Qwen: Qwen3.5-9B supports a 262,144 token context window and up to 262,144 output tokens.
What capabilities does Qwen: Qwen3.5-9B support?
Qwen: Qwen3.5-9B supports Streaming, Function calling, Vision, JSON mode, Playground.
Which providers offer Qwen: Qwen3.5-9B?
Qwen: Qwen3.5-9B is available from DeepInfra, Together.
How do providers handle data privacy for Qwen: Qwen3.5-9B?
2 of 2 providers report that customer data is not used for training, and 1 offer zero-data-retention routing. Retention, compliance, and privacy-policy links are listed per provider.

Browse all AI models