Qwen 3.5 Flash
Qwen 3.5 Flash is Alibaba Cloud's production-hosted multimodal model built on a hybrid linear-attention MoE architecture, offering a context window of 1M tokens and sub-second responsiveness for high-throughput agentic workloads.
- Input and output price
- Input $0.10, Output $0.40, Per 1M tokens
- 24h uptime
- Loading AI Gateway uptime
import { streamText } from 'ai'
const result = streamText({ model: 'alibaba/qwen3.5-flash', prompt: 'Why is the sky blue?'})Copy link to headingLatency24 hours
P50 time to first token (TTFT) on live AI Gateway traffic, in milliseconds. View the docs for more info.
Your use is subject to Alibaba Cloud's Terms & Privacy Policies.