Qwen: Qwen2.5 VL 32B Instruct

Name: Qwen: Qwen2.5 VL 32B Instruct
Brand: Qwen
SKU: qwen-qwen2.5-vl-32b-instruct
Price: 0.05 USD
Availability: InStock

byQwen

Qwen2.5-VL-32B is a multimodal vision-language model fine-tuned through reinforcement learning for enhanced mathematical reasoning, structured outputs, and visual problem-solving capabilities. It excels at visual analysis tasks, including object recognition, textual interpretation within images, and precise event localization in extended videos. Qwen2.5-VL-32B demonstrates state-of-the-art performance across multimodal benchmarks such as MMMU, MathVista, and VideoMME, while maintaining strong reasoning and clarity in text-based tasks like MMLU, mathematical problem-solving, and code generation.

Pricing

Input

$0.05 / 1M tokens

Output

$0.22 / 1M tokens

Specifications

Context Window16K tokens

Max Output16K tokens

Modalitymultimodal

Input Typestext, image

Output Typestext

Strategic Analysis 🔒

Unlock vCAIO insights to make better model decisions:

Governance Risk Rating (Low / Medium / High)
Quality Tier Classification
Best Use Cases & Tags
Strategic Verdict from vCAIO
AI-Verified Fit Scoring

Start Free Trial Sign In

Not sure if this model fits your use case?

Describe your task and get AI-verified recommendations in seconds.

Try Model Advisor

Popular model profiles

Pricing last updated: Invalid Date

Qwen: Qwen2.5 VL 32B Instruct

Pricing

Specifications

Strategic Analysis 🔒

Not sure if this model fits your use case?

Popular model profiles

Other Qwen Models

Qwen: Qwen VL Max

Qwen: Qwen-Max

Qwen: Qwen3 Max

Qwen2.5 72B Instruct