Skip to main content

ModelRegistry

Explore text, image, video and speech models

Find the right AI by model, modality and provider.

Browse models

Choose with facts

Turn the model directory into an actionable evaluation workflow.

Clear capabilities

Review modalities, protocols, limits and status.

Transparent coverage

See which providers expose each model.

Easy comparison

Compare candidates using consistent dimensions.

Start exploring the directory

Begin with a model, provider or organization.

Open model directory

Model spotlight

Qwen3 VL 30B A3B Instruct

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...

Compare models

Core capabilities

A compact view of the model contract before integration.

Input and output

See supported modalities and generation directions.

Limits and context

Check context, output limits and release details.

Provider coverage

Compare the providers and protocols that expose this version.

Qwen3 VL 30B A3B Instruct FAQ

Start building with Qwen3 VL 30B A3B Instruct

Move from model evaluation to a production-ready workflow.

Explore providers

Qwen3 VL 30B A3B Instruct

alibaba/qwen3-vl-30b-a3b-instruct

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...

Model family

Qwen

-

Focus areas

General-purpose modelsMultimodal AICoding

Core specifications

Released
2025-10-06
Context window
262K
Max output
33K
Model parameters
31B
Weights
Open

Modalities

Input
ImageText
Output
Text

Identity

Version
qwen3-vl-30b-a3b-instruct
Knowledge cutoff
2025-03-31
License
Apache-2.0

Resources

Capabilities

Open weightStructuredTemperatureTool call

Capabilities and fit

Model summary

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...

Text
Open weightStructuredTemperatureTool call
Output
Text
Capabilities
4
Catalog updated
2026-08-06T01:19:42.344Z
Status

Provider coverage

ProviderInputOutputPricing
OpenRouter
Qwen3 VL 30B A3B Instruct
Pricing unavailable
Kilo Gateway
Qwen: Qwen3 VL 30B A3B Instruct
Pricing unavailable
NovitaAI
qwen/qwen3-vl-30b-a3b-instruct
Pricing unavailable
Alibaba
Qwen3 VL 30B A3B Instruct
Pricing unavailable

FAQ