EN Submit a tool

qwen/qwen3-vl-8b-instruct

Rank #164 qwen

qwen/qwen3-vl-8b-instruct

Usage data date 2026-09-12 · updated daily

Usage and specs

Daily rank#164
Daily tokens10B
Daily requests10M
Daily change −32.4%
Released

What is qwen/qwen3-vl-8b-instruct

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...

qwen/qwen3-vl-8b-instruct capabilities

Input: Image、Text   Output: Text

JSON output Structured outputs Tool choice Tool use

qwen/qwen3-vl-8b-instruct API pricing

Input · per 1M tokens$0.12
Output · per 1M tokens$0.46

USD list-price snapshot, not independently verified; confirm current pricing and terms on the provider site before purchase.

Tools built on this model