QwenReasoningReasoningActive

Qwen: Qwen3 VL 32B Instruct

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

Specification
Model ID
qwen/qwen3-vl-32b-instruct
Modality
Text+image >text
Context
131.1K tokens
Input
$0.104/1M
Output
$0.416/1M
Updated
Sep 20, 2026
Scores
Capabilities
  • Supports tool calling
  • Supports vision inputs
  • Supports long-context workflows
  • Lighter reasoning profile