ALAlibaba/ Qwen
Qwen3 VL 32B Instruct
alibaba/qwen3-vl-32b-instruct
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...
Model information
Released
2025-10-23
Context window
131K
Max output
33K
Model parameters
33B
Weights
Open
Modalities
Input
ImageText
Output
Text
Identity
- Version
- qwen3-vl-32b-instruct
- License
- Apache-2.0
Resources
Capabilities
Open weightStructuredTemperatureTool call
Providers
| Provider | Provider model ID | Protocol | Context | Max output | Input | Output | Capabilities | Notes | Pricing | Streaming |
|---|---|---|---|---|---|---|---|---|---|---|
| OpenRouter | qwen/qwen3-vl-32b-instruct | openai_chat | 131K | 33K | AttachmentsTool callStructuredTemperature | Hugging Face: Qwen/Qwen3-VL-32B-Instruct | input $0.104/M · output $0.416/M | Yes | ||
| Kilo Gateway | qwen/qwen3-vl-32b-instruct | openai_chat | 131K | 33K | AttachmentsTool callStructuredTemperature | - | input $0.104/M · output $0.416/M | Yes | ||
| Alibaba | qwen3-vl-32b-instruct | custom | 131K | 33K | - | - | - | Yes |