Qwen3 VL 235B A22B Instruct
Qwen
qwen/qwen3-vl-235b-a22b-instructQwen3 VL 235B-A22B Instruct is Alibaba's large-scale multimodal MoE model, combining Qwen3's language capability with vision understanding for image, document, and video tasks. It suits demanding multimodal applications needing near-flagship quality.
Cost rate
0.4x
Context
262K
Released
Sep 23, 2025
Input
TextImage
Output
Text
Support
ReasoningTool callingStructured outputs
Best at
The categories where this model ranks highest.
Image understanding
#63 · top 43%
Business, Management, & Finance
#89 · top 23%
Legal & Government
#94 · top 26%
Mathematical
#98 · top 26%
Performance
Median latency and throughput measured across recent requests.
Throughput47 tok/s
Latency