Skip to content
Dashboard

Qwen 3.5 Flash

Qwen 3.5 Flash is Alibaba Cloud's production-hosted multimodal model built on a hybrid linear-attention MoE architecture, offering a context window of 1M tokens and sub-second responsiveness for high-throughput agentic workloads.

Input and output price
Input $0.10, Output $0.40, Per 1M tokens
24h uptime
Loading AI Gateway uptime
import { streamText } from 'ai'
const result = streamText({
model: 'alibaba/qwen3.5-flash',
prompt: 'Why is the sky blue?'
})
Read docs

Copy link to headingUptime

Direct request success rate on AI Gateway and per-provider. Visit the docs for more info.

Your use is subject to Alibaba Cloud's Terms & Privacy Policies.