EN Submit a tool

qwen/qwen3-vl-32b-instruct

Rank #117 qwen

qwen/qwen3-vl-32b-instruct

Usage data date 2026-09-12 · updated daily

Usage and specs

Daily rank#117
Daily tokens34.4B
Daily requests9.6M
Daily change +86.2%
Released

What is qwen/qwen3-vl-32b-instruct

Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...

qwen/qwen3-vl-32b-instruct capabilities

Input: Text、Image   Output: Text

JSON output Structured outputs Tool choice Tool use

qwen/qwen3-vl-32b-instruct API pricing

Input · per 1M tokens$0.10
Output · per 1M tokens$0.42

USD list-price snapshot, not independently verified; confirm current pricing and terms on the provider site before purchase.

Tools built on this model