Qwen: Qwen3 VL 30B A3B Instruct

Name: Qwen: Qwen3 VL 30B A3B Instruct
Brand: Qwen
SKU: qwen-qwen3-vl-30b-a3b-instruct
Price: 0.15 USD
Availability: InStock

byQwen

Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception of real-world/synthetic categories, 2D/3D spatial grounding, and long-form visual comprehension, achieving competitive multimodal benchmark results. For agentic use, it handles multi-image multi-turn instructions, video timeline alignments, GUI automation, and visual coding from sketches to debugged UI. Text performance matches flagship Qwen3 models, suiting document AI, OCR, UI assistance, spatial tasks, and agent research.

Pricing

Input

$0.15 / 1M tokens

Output

$0.60 / 1M tokens

Specifications

Context Window262K tokens

Max Output— tokens

Modalitymultimodal

Input Typestext, image

Output Typestext

Strategic Analysis 🔒

Unlock vCAIO insights to make better model decisions:

Governance Risk Rating (Low / Medium / High)
Quality Tier Classification
Best Use Cases & Tags
Strategic Verdict from vCAIO
AI-Verified Fit Scoring

Start Free Trial Sign In

Not sure if this model fits your use case?

Describe your task and get AI-verified recommendations in seconds.

Try Model Advisor

Popular model profiles

Pricing last updated: Invalid Date

Qwen: Qwen3 VL 30B A3B Instruct

Pricing

Specifications

Strategic Analysis 🔒

Not sure if this model fits your use case?

Popular model profiles

Other Qwen Models

Qwen: Qwen VL Max

Qwen: Qwen-Max

Qwen: Qwen3 Max

Qwen2.5 72B Instruct